from Salamence@mander.xyz to privacy@lemmy.ml on 10 Sep 02:47
https://mander.xyz/post/58014509
cross-posted from: hexbear.net/post/9497419
The Millennium Problems are a set of the most important open problems in mathematics. so far only 2 have been solved: the Poincare Conjecture by Grigori Perelman in 2010 and today OpenAI released this openai.com/index/navier-stokes-solution/
The only problem is, that their AI didn’t even come up with the solution itself. Mathematicians working on the same problem recently made big strides in solving the Navier Stokes equations, and their chat logs were scraped and used in training data before they could release the proof themselves.
Here is the unpaywalled statements of the Mathematician: mastodon.social/…/117233413705701198
Western AI companies have been tackling a bunch of open problems and techbro chuds cannot shut up about it, this is just Marketing and the fact they have to steal real people’s work to do so just proves it. They really want that IPO moneyyy <img alt="capitalist-laugh" src="https://hexbear.net/pictrs/image/fdc96d7d-9fc4-4fc0-95ef-17aaf25405a2.png">
threaded - newest
Everything AI does.os with stolen data. What makes this so special?
It’s like if they wrote gta VI in notepad and Microsoft released it saying they made it… so not really different except they didn’t even let the author publish it first
The fact that AI had solved this problem made big noise, and was used by many to legitimize AI as a source of progress, so setting the record straight on how it really did it is pretty important.
Beside, solving a millenial problem is a big deal. Even if it had been solved by human researchers and it was later found that they had plagiarized other human researchers, this would’ve been noteworthy. So I would say this is pretty special.
Woah. Navier Stokes has been solved? That’s a huge deal. And fuck openai for undercutting the people who made it possible.
Well “solved” is a bit of a weird way to say it. Was it proven as true, or was it disproven? Both of those change things in different ways.
Proven true that it is a stable formula, then that’s cool, and we know that issues with fluid simulations based off of that are a software and compute problem.
If it was disproven, then we learn precisely under what specific conditions nature becomes mathematically unpredictable, and then we can start developing different and better models.
My understanding is that it has only been shown to have closed form analytical solution for very simple cases. “Solved” in this case would mean doing the same for more complex geometries. It would mean finding new functions like spherical harmonics or the Bessel Function.
So sad that it has been solved this way. It feels stolen (even if it had not copied work)
Pay wall
Same here. Here’s a archive.org link: web.archive.org/…/openai-navier-stokes-math-disco…
This feels like when you’re telling a story or a joke and someone swoops in from the sidelines to steal the punchline. Only like 100x worse.
Key and Peele
Wait I thought someone in anthropic made some progress on this using Claude which was then leaked to openAI and then they mustered all their resources to do it faster and did it before.
x.com/konstiwohlwend/status/2097235335034056835
Steal their work? well, but mathematician worked a lot of years on the problem without a solution, means that their work wasn’t correct. Apart scientific papers are accessible by everyone, not only AIs. Maybe the mathematicians have to think a little more about teamwork, storing their work locally or encrypted, instead claims about copyright and AI stealing their work. But the capabillity of AIs is to be able to find interrelations between millons of datas in hours, something for what humans need decades to do the same. That is the point and also the danger of AI, it get the data it needs from everywhere and web content is open, AI don’t have any ethics or respect of copyrights, all web content is information as is. All you put online is public (or everything online) by definition, there isn’t any copyright to avoid it.
The original team had used AI to solve the problem. However, they did not want to publish AI slop (think one of the authors actually described their initial proof as such, which is hilarious in hindsight). Then, the not so open to anything but stealing, openai stole their almost finished work and used tons of compute to finish it.
Your reply is incoherent and out of touch with the reality of what happened.
This is interesting news but why is it in the privacy comm?
Why is the fact that an AI company is surveilling their users and stealing their work in a privacy comm?
surveilling users? chats are fed into the training data, I believe this is stated pretty clearly on every chat, something like "your chats are used to improve the model’
That’s right, surveilling.
Only in certain plans and you can opt out. I assume they were smart enough to opt out but on review I couldn’t actually find evidence of that so you may be right…
this is the whole selling-point of locallama and why local models are so rigorously sought after (despite seeming excessive(seemingly needless) spending on rigs )
I think this is the real problem here. Chats are being fed to AI. This is horendous. Do we know which chat system was raked by OpenAI.
Somewhere down the article:
I’m aware that the data is used to train models. Most people know this, it’s stated clearly. That’s not the story here. The story is IP theft, not privacy.
That’s the whole debate… If researchers used ChatGPT then they agreed to its terms and conditions, which are shitty and forfeiting privacy, IP and all that comes with it. All the major AI companies are not in the business for the sake or betterment of humanity, but to exacerbate the capitalist model in which you own nothing (funny how that’s how communism is sometimes described). I mean the whole drama with Altman and OpenAI, that basically killed the initial nonprofit, humanist goals of Open AI was very revealing. I read a quote the other day that said AI is a means for wealth to access skills and knowledge, while preventing those with skills and knowledge to access wealth… Couldn’t say it any better.
This is why use of AI needs to be a conscious use. We need to understand that everytime we use it, we are teaching it to improve itself, and if we are feeding it sensitive data, then we are being really stupid.
it would be a better fit for intellectualproperty community . or maybe security. i guess it sorta could be privacy. just not the way most people tend to focus on. intellectualproperty is the bigger issue. but it occured due to lack of privacy .
turns out if you throw hundreds of billions of dollars at math problems, you tend to find solutions, who knew?
the worst part, to me, is the implication that such solutions are a question of capability and not of resources
🎶plagerize Let no one’s else’s work evade your eyes So don’t shade your eyes But Plagerize Plagerize Plagerize 🎶 But of course don’t forget to call it… Research
these news site really need to research or write better.
No, Navier Stokes was not “solved”. We have been using numerical methods to solve these equations approximtely since forever. The exact analytical solutions that exist for special cases are also studied; most often, these are used as baseline for comparison against new theory , new experimental results and new computer simulations. So, we have been “solving” Navier Stokes.
What OpenAI manages to achieve is answering a very specific problem relate to Navier Stoke (sorry for hand waving, not a pure math major). It is something like: given an initial flow condition, does the Navier Stoke equation predict this flow field to blow up or it stays smooth for some finite time? Result: OpenAI manages to come up with a case where the solution blow up.
Mathematically, i bet some math geniuses can expand on this finding as like a “special property” of the Navier Stoke equations.
So no we are still far solving the whole equation analytically.
You mean that the solution didn’t converge? Like, what does “blow up” mean?
I mean, you kept something secret in Codex. A product purpose built to gather training data. What did you think would happen?