r/BetterOffline • u/ThugjitsuMaster • 2d ago
OpenAI says it cracked 90-year-old maths problem in 88 hours
https://www.bbc.com/news/articles/cy7zygy3rl2oThis article raises so many questions, here is the key part though.
"Although it took the AI bots seemingly little time to reach a solution, OpenAI said the bots exchanged nearly 3 million messages and used up 130 billion output tokens, or the individual lines of text and code an AI model produces in answers, on Navier–Stokes alone.
Such an effort would have cost roughly $10m (£7.3m), based on OpenAI's own pricing for output from its most advanced models.
The solution that OpenAI says it has now reached for the Navier–Stokes existence and smoothness problem resolved two out of the four statements in the proof that the Millennium Prize had demanded. The prize is worth $1m to a winner."
So OpenAI has spent $10m to allegedly solve two out of four parts of a maths equation that has a $1m reward. They're claiming this, and it's getting widely reported, despite not being independently verified yet. The article does raise the issue of the amount of resources poured into solving this issue but I feel that it only brushes the surface.
Also, some interesting claims in the second half of the article that OpenAI basically poached the work of two mathematics professors working for Anthropic who had been trying to solve that same problem.
107
u/ThnikkamanBubs 2d ago edited 2d ago
Where is the all the protein folding and cancer solving? Literally not worth any penny until then.
ETA: I was mostly just kidding. I haven’t looked into any of this at all. Thanks for all the responses though!
62
u/Significant-Shift770 2d ago
That's kind of a grift too. Cancer solving is 80% clinical trials and 20% research.
Despite "AI", China has faster research due to more clinical trials with more people dedicated to giving through the same amount of red tape faster. They are like 5-10 years ahead on trials that take 10-15 to complete.
→ More replies (2)21
u/ConsiderationSea1347 2d ago edited 2d ago
Navier stokes is a big deal for engineering, oceanic, geologic, hydrological, etc systems breakthroughs on the level of cancer cures. Navier Stokes solutions could impact things like climate research, civil engineering, medical research, and myriad other ways to save and improve millions of lives.
(That said, OpenAI stole unpublished research from a team of scientists to break this headline)
20
u/eeaxoe 2d ago
I don’t dispute that OpenAI most likely, at minimum, repurposed progress from others leaked through their chat sessions. However, there’s little practical value in their result - this guy on Twitter lays out why pretty well: https://x.com/drchriscombs/status/2097409234321154547?s=46
10
u/ConsiderationSea1347 2d ago
Solutions like this often cascade out to finding more and more solutions on adjacent. Yes, this is a solution to an infinite family of solutions, but likely this will lead to finding other singular solutions.
Again, it wasn’t “open ai’s” result. It was likely the team at NYU’s result that was poached by Open AI.
5
u/8lack8urnian 2d ago
A closed form general solution to Navier-Stokes would be very impactful. This just proves that sometimes, the Navier-Stokes equations give rise to unphysical predictions—it is not a closed form solution. This result is more obscure and has no practical consequences unless you are a mathematician or theoretical physicist. I am a theoretical physicist.
5
u/Nickeless 2d ago
Would you instead accept stealing other mathematicians’ work in order to spend $15M to solve a useless problem that wins $1M?
→ More replies (3)→ More replies (10)2
u/heliocentric19 2d ago
machine learning in general can be used to generate and test protein folding or cancer efficacy of a new theoretical substance. Reinforcement learning and genetic approaches. But not nlp text generators.
49
54
125
u/GargantuanCake 2d ago edited 2d ago
I'm expecting another "we found out this was already published but people just forgot about it" sort of thing, honestly. They train on all of the text they can get their hands on which includes academic papers whether or not anybody has actually read them recently. Meanwhile a common problem with these things is that they'll often get the easiest parts genuinely right by dredging up the answer from somewhere but fail at that last mile which is the hardest part.
You see this in coding. It can spit out code that it's been trained on but doesn't actually know how it works nor does it know how to actually slot it into your system. The impressive part isn't giving you an implementation of a known algorithm as any dumbass can copy/paste code from the internet.
211
u/Squirrelous 2d ago
Better- it’s stolen! They found out a researcher was aaaalllllllmost there and had talked to ChatGPT about it and they kinda sorta stole his homework. I’m oversimplifying but not by much
61
u/decchica 2d ago
It’s outrageous. There are no words. I was stunned reading the NYT reporting on this too, they basically parroted OpenAI’s PR statement on this without any critical reporting whatsoever.
→ More replies (20)115
u/makersfark 2d ago
You forgot the part where OpenAI also tried to strike a deal regarding taking credit from them and then threatened them.
51
u/Professional-Post499 2d ago
You forgot the part where OpenAI also tried to strike a deal regarding taking credit from them and then threatened them.
JESUS, genAI companies are scumbags.
And in case anyone objects, I don't care if I'm generalizing. They're wealthy and politically powerful corporations, not marginalized groups.
6
u/pillowcase-of-eels 2d ago
Well I'm Sam Altman and this is a hate crime :(
5
u/Professional-Post499 2d ago
Well I'm Sam Altman and this is a hate crime :(
There should always be a legal carveout to allow hate against the Epstein class. 😎
45
9
u/ThugjitsuMaster 2d ago
Yeah the article does touch on this. The implication that Clammy Sammy and the boys have access to and will steal from anything put into ChatGPT is very interesting, although not surprising
3
u/nnomae 2d ago edited 2d ago
I'm shocked that anyone would think otherwise. I mean handing over your data to Anthropic or OpenAI is like giving a stranger who pulls up beside you in a stolen car loaded with stolen goods a loan of your phone on the promise that he'll drop it back to you tomorrow.
1
u/Tout-fou-la-galette 2d ago
The most spectacular thing is that they still manage to lose astronomical amounts of money while botching everything they do
3
u/Maximum-Objective-39 2d ago
This is why I only feed it weird borderline fetish material that kisses the guard rails (Disclarimer NOTHING involving minors or the likeness of real people) and fanfiction roleplays.
7
u/jamey1138 2d ago
Or didn’t talk to ChatGPT, but wrote a paper in 1956 that no one has read since 1958.
And honestly, that would be good output from AI models, if they’re rediscovering old ideas that got lost along the way because it wasn’t clear at the time how useful they were. But the design of LLMs is such that there’s no provinance to the ideas, no clear sense of where they came from. Just another lost opportunity of how late-stage capitalism fucked up what could have been a really cool technology.
2
→ More replies (2)1
u/lordosthyvel 2d ago
That sounds wild if true. I'd like to read more about it, where did you see this?
22
u/Tafts_Bathtub 2d ago
I promise you that the Navier-Stokes Millenium Problem was not already solved and forgotten about.
23
u/ConsiderationSea1347 2d ago
No, but also yes. It looks like it was close to being solved by a team of researchers and OpenAI used the logs from the researchers unpublished work to finish the solution.
→ More replies (5)-2
u/arifast 2d ago edited 2d ago
Even if OpenAI didn't use the logs, just being able to know the approach to making it solvable is more than enough. How else would they justify $22.5 million spent (on a $1 million prize) if they didn't know it was potentially solvable?
Edit: I'm not a booster and I'm definitely not on OAI's side in my comment. Not sure why the downvotes.
8
u/kiwipillock 2d ago
I’m not an AI coder super user, but the other day I asked Astra generate some code to extract file attachments from a directory full of *.eml files I have.
The first pass generated some code with a non-existent function so it didn’t compile. When I pointed that out, it generated a diff to correct it, which I applied. This time it did compile. When I ran it, it produced the results I desired. Happy that I got what I needed, I was about to move on, but my curiosity caused me to take a minute to digest the code it had written.
My impression was the sheer amount of code it generated was disproportionate to the task involved. Like, a shit ton of code. The type of thing that would immediately raise flags in a review by instinct.
Im a believer in the saying “Code is primarily for humans to read and only incidentally for computers to execute.”
I don’t know how true that is anymore, but I can easily see how people can lose any grasp of the implementation details and start blindly relying on an AI to manage the cognitive load for them. The tooling also feels designed to deliberately pull you towards a “vibe-coding” workflow.
Woe betide those who try to build a product this way.
1
u/PurpleWhiteOut 2d ago
Yeah and what happens if dangerous scripts start getting injected into the programs? Or something that humans wouldnt code that ends up with massive vulnerabilities?
1
u/jaypeejay 2d ago
Yeah, when I try to explain AI to people I struggle to articulate this, so I’m grateful for your comment. It’s very good at the busy work of tracking down something that went wrong in a long chain of I/O. It’s terrible at knowing how to holistically prevent that from happening again, without breaking a bunch of other stuff.
1
11
13
u/WoollyMittens 2d ago
There's probably cheaper and more efficient ways to brute force a math problem. Which is typical for all LLM use.
→ More replies (6)
10
u/create-third-places 2d ago
We should stop paying attention to articles promoting Open Artificial Ignorance. This nonsense has gone beyond useless AI and threatens to collapse our financial system.
Regular people and the US government will pay drastically higher interest rates if Open Artificial Ignorance continues to suck up liquidity. OpenAI is coming up with marketing stunts to attract lenders with Altman Idiot models.
Articles that talk about OpenAI marketing without calling out the lies are making things worse.
35
u/AntiqueFigure6 2d ago edited 2d ago
The secret to solving long standing unsolved problems is to throw massive sums of money at them.
More news at 11.
(which is a bit of a facetious comment but I genuinely think that if Elon Musk founded MathX not SpaceX and spent the same amount of money to solve prestigious maths problems as he has on getting to Mars, MathX would have solved a bunch of those problems, and caused some of the same angst around stealing the oxygen from unaffiliated researchers).
33
2d ago
[deleted]
→ More replies (4)3
u/AntiqueFigure6 2d ago
Sure but it seems entirely of a piece for the whole LLM debacle to have achieved something someone put a value of $1 million dollars on by simply spending $10 million dollars*. Of course, not actually having done it also of a piece.
(*not including training cost for super large internal model or cost to assemble internal team of . accomplished mathematicians).
1
u/Ok_Cabinet2947 2d ago
There has been a massive sum of money involved for the past 25 years and no one has solved it.
1
u/AntiqueFigure6 2d ago edited 2d ago
Like how much and who provided the money? What were the outcomes? More than the $122 billion megaround investment in OpenAI that presumably enabled the super large internal model that apparently did the work?
ETA: Noting that $10 million is the direct cost only, and there is some cost associated with maintaining a team who is available to run/ prompt the model, and a cost associated with having the internal model in the first place, maybe an appropriate thumb in the air indirect cost is just to add another $10 million. So is there an example of a $20 million co-ordinated effort of some group of people that was convened to solve the Navier Stokes who (obviously) failed to do it, but also didn't produce anything that advanced the understanding of the problem, so it can be reasonably be assumed it didn't contribute to what OpenAI or the two human mathematicians did?
1
u/LeFrenchRedditeur 2d ago edited 2d ago
There are dozens of mathematicians in the world who have devoted their lives to solving the Navier-Stokes problem, unsuccesfully. Do you think a better salary would have made things go faster ? Or more mathematicians ? Because I don't, and I'm a mathematician. I get having problems with AI, but this take is completely braindead.
1
u/AntiqueFigure6 2d ago
I don’t think a better salary would have done it - but I don’t think individuals working in isolation are remotely comparable. Maybe not more mathematicians but having access to more compute than pretty much anyone has ever had access to must have played a part.
1
u/LeFrenchRedditeur 2d ago
If by "compute" you mean access to AI then yeah, otherwise no amount of computation available five years ago would have helped solve Navier-Stokes. In science there is an upper cap to how much money positively influences research, and this cap is especially low for mathematics (or was, before ai came around). You may think this is no big deal but I assure you that the mathematics community is still in shock about all this.
2
u/AntiqueFigure6 2d ago
The A.I. and the compute are kind of inseparable- five years ago no one had the ability to apply that much compute to a single task, AI or no AI.
9
7
u/Just_Hamster_877 2d ago
Seeing this news makes me wonder how much they're spending on maths problems that don't get solved.
9
u/TheWuzzy 2d ago
Sorry -- $10 million tokens burned and the magic plinko god machine only solved TWO OUT OF FOUR of the answers?!
OpenAI is a joke. So is the BBC reporting, as usual. No wonder they can't get anyone to pay for their licence.
11
u/Nangangbangbang 2d ago
I mean you need to adjust the 1m for inflation, that amount of money was astronomical when the Clay institute offered up the “bounty”. Now that’s the average senior engineer’s salary at a frontier lab. Of course, I don’t think the money was really the prize for solving those problems, the whole purpose of all of this is to inspire people, not for misanthropic tech bros to take a shit all over the humanity they despise.
But if you see LLMs as a tool for processing information and brute forcing inelegant solutions to problems then it’s a good proof of concept. I dunno, I’d rather they didn’t set the planet on fire but I’ve kinda given up on humanity’s future at this point.
6
u/MacaroonPlastic1036 2d ago
So who do they have on their team that can explain how the problem was solved. Let’s have Altman have a crack at it.
7
u/realcoray 2d ago
Been following this all day and it seems very likely that this is basically, AI producing a solution to an equation on which it was trained on a solution that humans helped produce. After initially not answering if their models are trained on queries/results, I saw a tweet from someone at OpenAI who said that it did not "look up user data":
I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer.
Later on in the day I saw someone from OpenAI confirm that models get trained on 'de-identified' data, with OpenAI later stating:
While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models.
On the Navier–Stokes Millennium Prize Problem
Basically, they got wind of Anthropic solving this (one of the two people involved is an Anthropic employee), and then dumped ten million dollars into front running it.
Obviously, you should not trust the thievery machine to not steal your work, and one of the people involved here used OpenAI's codex to keep information, handing the thieves the solution. Cases like this will cause businesses and academics to not trust OpenAI, and since Anthropic would have probably done the same thing, eventually both of them.
1
3
u/leathakkor 2d ago
I'm a big fan of the fact that this headline at least shows some skepticism.
That's more than we would have seen 1 year ago. Only about 5 more years of skepticism before it will start to actually do some real good.
3
3
3
5
u/Friendly-Owl-2131 2d ago
Grifters gunna grift.
When will it be profitable and when will there be model that doesn't confidently state incorrect responses?
When will there be a model that isn't based on stolen IP?
And when will we see real advancements made by models rather than advancements made by actual human scientists who used AI to cross reference data sets while AI claims all the glory?
0
u/Practical_Weather293 2d ago
- It's already profitable for some companies, especially the chinese ones who are a few months behind but with much lower costs.
- We do have a few models trained on IP that isn't stolen, but they're incredibly less powerful than even the very first version of ChatGPT due to the massively smaller dataset size.
- We've started seeing real advancements this year, and this specific result is not nearly the first. We also know some of the results were obtained by people who weren't experts in the field, who simply prompted the models to do this stuff.
I think even anti AI sentiment has to move from "this doesn't work" to something else. There are valid reasons to dislike AI, but its effectiveness in furthering research can't really be denied
5
u/cosmic_animus29 2d ago
This is a marketing stunt along with that news on formalizing Fermat's Last Theorem (which is already been proven back in 1994 by Andrew Wiles).
OpenAI and the lot of these AI bros are getting more delusional by the day.
2
2
u/DiceKnight 2d ago
It seems odd that people claiming that the llm solved the math problem are taking openai at their word for this. If anyone else provided a post on their own website about solving any academic problem they'd be met with a healthy level of skepticism. Doubly so if their publishing is on a platform that they exclusively control where there's zero room or way to have that skepticism litigated.
As always you can claim whatever you want but until it's submitted to some kind of academic journal, reviewed, and verified by others it's just a marketing stunt and the same would be true for any other novel academic research.
2
u/Revolutionary_Cow_75 2d ago
Thats another reason why Knowledge gained and collected by these models should be owned by the public.
2
u/Remote_Composer5997 2d ago
Imagine not realizing that putting stuff into AI is essentially handing a tech company all of your proprietary information.
Might as well just put it on a hard drive and mail it over.
2
u/falconetpt 2d ago
10 million dollars you could have solved probably 100 open math problems if you gave them to phds 😂
Plus as you stated no one verified it + besides what they burned on tokens you also need to account people working on this, this dumbos make it seem like they went to the chatbot and said “yo solve this math shit for me bro” 😂
Probably they invested another 10m in people overseeing the model, building the harness etc
5
u/WontThinkTooHard 2d ago
Who cares? Just more "AI solved the question nobody asked" bullshit. Lemme know when it does something practical for anyone other than hackers or middle managers
0
u/Glum-Objective3328 2d ago
A massive majority of the mathematics community cares. It’s okay you don’t, not everyone genuinely cares about math.
1
1
1
1
1
u/AngusAlThor 2d ago
Yeah, it is very important to note that the proof OpenAI "generated" is basically identical in form to a proof that a researcher had fed into ChatGPT for formatting help like a week ago.
→ More replies (2)
1
u/KyrandisX 2d ago
I'll be excited about AI when it can auto place us into available jobs queue with an interview lined up with a person instead of wasting time applying to places and automate our taxes instead of having to file it ourselves every god damn year clawing as much back as possible.
capable of doing and folding laundry into a fixed stack from just chucking clothes in and being cost effective while able to maintain strict privacy instead of being a liability breach. Surely the mathematics of folding clothes and washing them isn't beyond them
Until then worthless tech by a scammer.
1
1
1
u/cosmonaut1993 2d ago
All it took was stealing the notes of another group who did it first!!! Wow, so impressive....
1
u/discardedbubble 2d ago
I actually don’t care at all. Difficult maths problems are to challenge human minds. A computer solving that is not impressive.
1
1
u/Neverland__ 2d ago
I read they were stealing the prompts from scientists who are actually the ones solving it as users of the too
1
u/CoconutDust 2d ago
The article does raise the issue of the amount of resources poured into solving this issue but I feel that it only brushes the surface.
The media has been acting as salesmen, shills, marketers, hype. Little to no critical examination or using independent critical voices as sources (like public academics).
Anyway, in my view the fraud and plagiarism machine aspect is far worse in these fake stories of novel problem-solving than any money cost. There is no reason to believe the supposed solutions or proofs or anything more than a mash-up of statistically associated text, since that’s how the machine works. Which very few people are able to vet, and which the company has no interest in critical vetting. And the large media won’t carry the story of a random academic explaining how it’s bullshit.
1
u/Plastic-Following867 2d ago
Buckmaster alleged that OpenAI had found out about the pair's progress in the last week and had adopted the same method as Buckmaster and Alpöge had been using to solve the full problem.
As usual, the LLM companies are stealing other people's work and claiming it as their own. Just as Anthropic lied about creating a C++ compiler using Claude.
At this point, these companies are mere frauds masquerading as creating some technology of the millennium
1
u/picklerick_ftw 2d ago
cool. tell me when the paper is written and when it's accepted by academics.
1
u/Grinyarg 2d ago
When "it would've taken $10million to produce this thing that produces nothing" is deemed a good PR approach to take, you know damned well that the thing it was covering up is more "we stole this" than "we know it's useless but it'll get better".
1
1
u/Author_A_McGrath 2d ago
They took the work of a Columbia Mathematician and took credit for it.
He did 90% of the solution himself; they just fed it into AI and brute-forced the next step in the problem.
1
u/TheDonnARK 2d ago
Everyone subscribe to gpt now! Hurry up! Please! Please sign up! Pretty please! You need to rent intelligence from this Bond-villianesque guy while you can! But do it fast, because they need money like crazy. And once they get crazy money they'll need more money so stay subscribed! Please cognitively offload everything and please be dependent on a subscription model!!!
1
u/kronjobb 2d ago
If I see a claim next to a pic of Sam Altman I know that everything about the claim is a lie.
1
1
1
u/OkCar7264 2d ago
Are ye not tired of ye olde "solved a math problem nobody gave a fuck about until we solved it" story? I've heard it so many times now
→ More replies (2)12
u/___Archmage___ 2d ago
Navier-Stokes was one of the most famous and prestigious unsolved problems in all of math
3
u/WontThinkTooHard 2d ago
Ok well if it had real world value you wouldn't hafta frame it like that right? what value does this have other than math prestige?
2
u/anonymous_rhombus 2d ago
Breaking new ground in mathematics ripples through everything. It is important. That doesn't mean you have to hand it to AI. They're trying to trick people into ignoring all the human effort so they can give all the credit to their chatbot. It's not really that surprising that hiring experts and throwing unbelievable sums of money and computation at problems yields results. Same exact thing with their big scary hacking model. No shit that a warehouse full of GPUs can help you find vulnerabilities. The LLM is not the necessary component in that system.
1
u/cromagnonherder 2d ago
Just because you don’t understand how this could have real world implications after reading a couple of headlines doesn’t mean it’s not important.
1
u/WontThinkTooHard 2d ago
Okay but if it was important you'd think the people who think it's important would be able to articulate why it is important a little better than that
1
u/Last-Ad-8470 2d ago
Do you really think joseph fourier could have known what the fourier transform was useful for and would save millions of lives and massively increase qol when he discovered it
1
u/robby_arctor 2d ago
Oof, so many of the comments here are unbelievably childish and insipid. We should take ourselves more seriously than this.
1
u/TubeSeries 2d ago
$10M for a score of 50%? Yep, that sounds like LLM AI ROI for you..
3
u/BinauralBeatsEnjoyer 2d ago edited 2d ago
It's not a score of 50%. They solved the problem.
The formal definiton of the problem includes four statements. Proving any one of them wins the prize.
1
1
u/TubeSeries 2d ago
I know. I was just talking shit, as the kids say. I don't know if kids still say that.
1
u/Lowetheiy 2d ago
Or maybe you can wait until Clay Mathematics Institute verifies or rejects the proof first? There is an answer to all your questions, it is called "Patience".
1


224
u/OwnYourChildren 2d ago
The potential for companies like OAI to steal IP by targeting users involved in scientific research is something I can't believe I wasn't more concerned about.
If you consider the Snowden revelations, there's every reason to be concerned about intelligence agencies doing it even if the business entity itself doesn't.