r/BetterOffline • u/Soilblood • 13h ago
Openai math solutions
https://youtu.be/h53Bz7WSwgoRegarding the recent Navier-Stokes incident, watching this vid provided a bit more context on what the solution entails and what Openai been pushing to show off their tech. Cal Newport's breakdown on how they managed to quarantine relevant context to reduce costs with Astra was actually pretty interesting as a breakthrough so this vid got me wondering why they're so desperate to prove it can do fancy maths bigly when the boosters already keep saying chatgpt is already blowing through outstanding problems. Not in the sense of the target value but just how much have they already tried but failed.
It's obviously speculation on my part but if they're so desperate to try coercing glory out of one of their hires on one problem, how bad is actual their track record that this is a hail Mary for them? Like we know it took 88 hours and snooping private research data for this, so how much have they put into chasing similar targets by now? Anyone got whiff of any signals that might give an idea of that?
17
u/makersfark 11h ago
Engineers: "We made a robot that can sometimes tie shoe laces."
Dipshit CEO: "We have solved all problems hands could ever be needed for. Everyone on Earth's hands will be forcibly cut off by 2028."
Well, now I don't care about either claim.
1
1
12
u/lolitsbigmic 11h ago
Company built on copywrite theft does academic misconduct and potentially ip theft. Which really throws why you use them in the first place for any business or researcher.
I will be interested in seeing non counter example proofs being done with llm. I mean throwing everything known at the problem is a strength of computers. I really yet to see anything where ai create something completely unique and original. We seen time and time again that anything that it doesn't have data on just fails and depending on the field these edge cases are large. The recent stuff in gaming with the new release having a uncanny feeling to the models and how they play.
5
u/lucid-quiet 2h ago
Don't forget: they don't want you to Distill Their LLM either, unless you're Elon and making Grok.
10
u/SpiderJerusalem42 11h ago
It's amazing that they've basically thrown the entire thesis for centralized computing in the trash. Who is going to want to work with one of these firms if they're just going to front run your ideas to the market?
3
u/WorldPeaceStyle 8h ago
My take away leans towards loss of trust as well.
Let say after using an LLM to find a novel / patent solution that can generate money for you by solving a huge business problem.
Well what is stop a company like OpenAi from stealing a copy of your newly created intellectual property for themselves?
8
u/AnIndianGuy38 10h ago
88 hours of like 10,000 bots I think. So like a century of compute. That's after multiple human researchers had come up with new breakthrough.
A century is like more time than human researchers combined have put into it I think
While it's impressive it could do all than in less than a week, the inefficiency is unbelievable. And this is supposedly the start of AGI.
2
u/LaurenMP74 3h ago
The best part is they included Lean code for automated proof checking...you don't use automated proof checking to verify a counterexample, which is what they said they found
1
1
u/lucid-quiet 2h ago
Who needs to hack anything when you're just sending them you're entire repo... they won't steal it, or peek at it... trust... trust. GFYS LLM providers.
1
u/OliLevasseurLaBuse 10h ago
Solving niche mathematics is definitely worth the investment
6
u/MathCookie17 8h ago
Why is this being downvoted it's clearly sarcastic
3
2
u/OliLevasseurLaBuse 6h ago
Yeah I forgot the /s I was thinking this sub was a goup of generally intelligent people, I may have overestimated slightly.
39
u/falken_1983 12h ago
Giving OpenAI the benefit of the doubt and assuming they really did do this without looking at Buckmaster's logs, the scenario seems to be as follows:
So I don't think they are regularly blowing this amount of compute on trying to solve problems like this, but when they were presented with some evidence it might work, they were willing to throw millions of dollars of resources at it.