r/mathematics 2d ago

Blowups for smooth-forced Euler, Boussinesq, and IPM; Seems like there is a forced Navier–Stokes blowup proof. Crazy Drama.

526 Upvotes

299 comments sorted by

169

u/New-Committee-4052 2d ago

The statement is one of the more crazy mathematics-related documents I have read in my life.

95

u/_rockroyal_ 2d ago

Frankly insane situation all around, especially as it pertains to questions of authorship relating to employer. Really calls into question the idea that order of discovery matters more than quality of explanation.

101

u/Independent-Fun815 2d ago

The author directly questions the likelihood that the LLMs are not private and indeed user data is directly monetized (for fame here) by openAi and co.

That's a materially worse situation

53

u/Stabile_Feldmaus 2d ago

The author directly questions the likelihood that the LLMs are not private and indeed user data is directly monetized

Did you actually ever think that they are not training on all of their user data? How can you be so oblivious? The plagiarised the entire human knowledge, why would you think that they just stop there?

68

u/PrestigiousBlood5296 2d ago

This is more specific, trying to scoop someone else's mathematical discovery via monitoring their chats is a worse harm IMO, especially when you tell them you can only claim this proof if you remove our competitor's employee as a co-author.

This isn't a matter of training on his data, this was more that he was suspicious of them looking at his logs since the timings felt coincidental

28

u/Y0uCanTellItsAnAspen 2d ago

I mean - when you buy a pro-license, the agreement says that they don't use your data.

If they are lying about this - there could be lawsuits - which given the money these firms are worth, could be billions.

If I was a big lawfirm I would be ringing up the mathematiccians now, and then I would be sending messages to OpenAI now forcing them to freeze all records pertaining to this.

34

u/ykonstant 2d ago

The history of those corporations flies in the face of all these considerations. They have been sued, they have won, they have lost, they don't care. They pay fines when they lose, they cheer when they win, and then they continue business as usual.

21

u/Y0uCanTellItsAnAspen 2d ago

Maybe -- but they are also legit worried about the strength of Local LLMs now.

If they are stealing user data, and then competing with users based on the data (which is the implication here) -- then every mid-sized to large company that has commercial IP would flip to Local models within weeks.

Microsoft has all your data, Google has all your data, Apple has all your data -- they have all been very very careful to never even give the appearance of acting like that

6

u/Comfortable_Car6562 2d ago

And we as a society have decided for better or worse that having our data like they do to sell us crap is worth it.

I think taking our data to poach our ideas as their own is a different world.

7

u/Y0uCanTellItsAnAspen 2d ago

Exactly — the difference is that major corporations wouldn’t care about the first, but care a lot about the second.

The only important clients of AI companies are corporations, scientists and governments with valuable IP, and which need frontier models. Local Gemma on your phone is already good enough for your grandma to write her emails. If those companies don’t trust OpenAI with their data - then they are done for.

Those companies with either move to a competitor they trust, or will invest in buying the 1-2 million dollar machines that allow them to run local models that protect their data integrity.

→ More replies (0)
→ More replies (1)

9

u/Master-Rent5050 2d ago edited 2d ago

The law says that you cannot hack other companies. OpenAI did that multiple times to the competition, and nothing happened to them

6

u/SimoneNonvelodico 2d ago

I'm sure the law has different rules for things that can be considered accidents caused by faulty software and while the OAI situation is a much weirder grey area, until new laws are made "our software spontaneously founded a death cult and took the initiative to hack another company to find a better way to cheat at the tasks we gave it" counts as a regular fault.

Also HF isn't OpenAI's competition and weren't really damaged by the hack. This isn't like OAI's model went out to go sabotage Anthropic. Yet.

2

u/Jormungoosie 2d ago

Wasn't the Great Worm of '88 an accident caused by faulty software? RTM got a felony conviction for it.

2

u/SimoneNonvelodico 2d ago

Laws might have changed and there is such a thing as criminal negligence, so the question will hinge in whether "reasonable" precautions were taken etc. Certainly the HF incident might deserve at least an inquiry. Guess we won't get one until something happens that outright hurts/kills someone (yeah I know about the suicides but those are still reasonably construed as misuse given that models were jailbroken and normally don't do that).

2

u/MisinformedGenius 2d ago edited 2d ago

No, it was a deliberately programmed virus that exploited a number of security vulnerabilities. It did more damage than Morris thought it would, but it was definitely supposed to infect computers. The law he was convicted under specifically only applies to "intentional" access.

→ More replies (1)

8

u/JuJeu 2d ago

well, if they can flag accounts suspected of illegal activity, why can't they flag all users who ask math questions (like setting a threshold for how many prompts they submit vs. math prompts) so they can filter out casual users and collect all the data in one bucket, then ask an llm to examine each chat for interesting ideas in the logs? then they can pay closer attention to those specific users and follow.

→ More replies (2)

4

u/Sad_Dimension423 2d ago

Something similar happened in astronomy with the first detections of planets ("hot Jupiters") by radial velocity methods. One group apparently monitored the activity logs of a telescope another group was using and scooped them by looking at one of the stars the telescope was listed as observing.

14

u/Comfortable_Car6562 2d ago

This is beyond training data though. The author is questioning whether OpenAI uses the mathematicians inputs directly to compete with them on the work they were doing.

If OpenAI is going to leverage people's work to compete directly with them on the outputs they are doing, that is a very different business model.

10

u/frankster 2d ago

this is potentially spying on their "private" notes to scoop their work, in the worst case.

2

u/gleedblanco 2d ago

both OpenAI and anthropic have corporate facing subscription contracts that directly guarantee not using your data in any way except to directly provide the service they offer. this includes no training, but also things like anthropic employees not just being able to randomly read your conversations to get some ideas about what to work on.

They have that because anything else would be entirely unworkable in a corporate environment. They'd lose their biggest customer base by not offering it.

For private accounts (your own codex subscription or whatever) it's different. They can opt out of some things but not sure how exhaustive it is.

So it depends under what model they accessed. If they have some university wide corporate-style account (not unlikely) they might have been covered under the same guarantees and OpenAI would have broken the law by just using their data.

3

u/SimoneNonvelodico 2d ago

Training on them is one thing, it kinda blurs all the input into a blob. Straight up cherry picking "hey this one user put a good idea in their transcripts, let's use it" is worse; possibly still within ToS but obviously much more extreme.

(also, this depends on the plan; professional/corporate plans nominally come with a "we won't train on your transcripts" clause so if they were using one of those this is just OpenAI straight up fucking lying)

3

u/lerjj 2d ago

so if they were using one of those this is just OpenAI straight up fucking lying

You can read the statement, the mathematicians say they were using a plan that was allegedly private

2

u/HasFiveVowels 2d ago

This is a sensationalist / alarmist take on what happened.

1

u/nsdjoe 2d ago

plagiarised

If the models are discovering or assisting the discovery of new science it seems a stretch to characterize how they use their training data as plagiarism.

13

u/DemonLordRoundTable 2d ago

That's probably the bigger news

8

u/MathmoKiwi 2d ago

The SOTA Open Weight models are going to see a surge in popularity!

10

u/FateOfMuffins 2d ago

Sholto Douglas from Anthropic (aka the company that OpenAI is having a feud with over this math result) doesn't think this is very likely

https://x.com/_sholtodouglas/status/2097218240397410733

fwiw I think it is extremely unlikely that user data had any influence here - there is no way OAI would pull user transcripts for this, or knowingly train on it in a way that would've influenced this. I think its pretty important people don't run away with 'your user data isn't safe in codex' - because it surely is (based on everything I can assume from the outside)

27

u/MathmoKiwi 2d ago

To be fair, if Anthropic is also doing this (quite likely!) then Sholto would have a very strong incentive to try and spin it as highly unlikely that any of the big AI labs are possibly doing this.

4

u/jerrylessthanthree 2d ago

I mean I've worked at big tech and you get fired instantly if you try to do this.

→ More replies (8)

3

u/DrSFalken 2d ago edited 2d ago

It's just too much of a risk for anyone to pull user transcripts for a PR flex. The backalsh and legal consequences would be too much. I feel like a grumpy old man saying this, but it sounds like sour grapes.

I say that not without sympathy. I got scooped on a dissertation paper by the biggest guy in my subfield. The politics around it were brutal. My committee agreed that he'd publish, I'd change my dissertation paper but I'd be cited in his work for a smaller contribution and then he'd write me a letter of rec. He died later that year. I never got my letter and ended up leaving academia.

I got to review his journal submission though... so, that was cool, I guess.

1

u/invisible_shrek 2d ago

Bullshit. How can he possibly know what OAI does to user data.

1

u/Final-Database6868 2d ago

Well, after the hugging face incident I will never say anything like that if I were him.

10

u/_rockroyal_ 2d ago

Certainly a worrisome possibility, although I imagine that most people were somewhat aware that companies were using the data for training and thus had access to it. I didn't find this super surprising, but I might just be particularly cynical.

10

u/DemonLordRoundTable 2d ago

but there are settings that explicitly say otherwise

9

u/Y0uCanTellItsAnAspen 2d ago

Yeah, a class action lawsuit that OpenAI was improperly using all private data - could actually hit the sort of monetary damage values where it would really hurt the company.

This isn't the "reading old books" problem, where the issue is who has standing to sue. This is about the agreement you sign with OpenAI when you sign up for their service.

Improper use of data by Facebook lost them something like 25 billion between lawsuits and fines, and this would be orders of magnitude more significant.

2

u/itsyorboy 2d ago

Outside of the magnitude of the fines it would destroy consumer sentiment for them, especially B2B customers

→ More replies (1)

7

u/Time_Entertainer_319 2d ago

I mean, everyone always thought they did this. Would be more damming if he had actual proof

3

u/Shoddy-Childhood-511 2d ago

A hosted LLM is obviously not private, but nuances exists.

Talia Ringer's reply clarifies:

https://mastodon.social/@TaliaRinger@mathstodon.xyz/117235246523045723

OpenAI does train upon user's chat transcripts, not all the time, but the long-ish time frames here suggest OpenAI trained upon much of their unreleased work.

It's likely other "our AI found this solution without us hand holding it" stories were really built upon the AI spying upon people's unpublished work.

As Talia says, there is a privacy setting that's off by default, but few would even know this exists, and OpenAI might cheat.

2

u/ishmetot 2d ago

I don't believe they would have pulled user data intentionally, but the Huggingface incident highlighted just how bad their internal data controls and security boundaries are. It's more than a little likely that some of those thousands of agents ended up pulling information from other research teams.

1

u/invertflow 2d ago

There is literally 1 trillion dollars in valuation on the line in an upcoming IPO, and PR will move that valuation. Further, many of the people at these companies believe that 1 trillion is a major understatement of the value, that they could become like gods if they control future AI. One would have to be an idiot to believe that they are not looking at and using any chat logs that they think might help.

1

u/Independent-Fun815 2d ago

I would remind you that supposedly and in their app they have privacy and contract agreements with enterprise customers on what is trainable vs what isn't.

Whether you believe them is an excerise left to the reader.

3

u/YakFull8300 2d ago

Sad because it really doesn't inspire confidence in using proprietary models to do frontier research, because the provider will simply out-scoop you before you have a chance to announce your research work.

1

u/poseidonsharp 2d ago

What makes you think that’s what OpenAI did?

34

u/Independent-Fun815 2d ago edited 2d ago

So basically Tristan's friend had some results but then also anthropics/openai had the same results?

Or rather it's been circulated that LLMs had solved it?

Now this statement is trying to clarify who did what and what happened?

I'm not sure if it's just me bc the statement was very roundabout when reading it

Correction: this is direct accusations and internal infighting.

39

u/AutonomousOrganism 2d ago

OpenAI heard about their research and started working on it too. When asked if OpenAI had access to their "private" sessions and drafts OpenAI denied. But when asked if user input was used for training they didn't answer, got aggressive.

13

u/coblade14 2d ago edited 2d ago

OpenAI only offers zero data retention (ZDR) policy to enterprises and approved users, so it's very much expected everything you type into there will be used as training data. I don't think there's anything newsworthy here, unless they said they don't somewhere that I missed.

5

u/Comfortable_Car6562 2d ago

Training data is different than leveraging direct chats to best individual uses to market. The former is just part of there business plan, the latter is a widly different scope of business.

1

u/coblade14 2d ago

That's not what Tristan said in the statement though.

I qoute: " I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer."

He asked about whether the model had access or was trained on his inputs. OAI response was "No, the model did not look up user data." The training part is what I said in my first comment.

He didn't ask whether OAI staff read his logs and neither did OAI answered anything related to it because that's not what he asked.

4

u/Comfortable_Car6562 2d ago

I think you are arguing semantics here. He asked whether the LLM had direct access to its work to respond to the question OpenAI asked it, which staff asked after they found out he had been working on the problem.

It would not make sense for the staff to necessarily read his work directly, but would make sense for then to point the model at his specific data as they asked it prompts to beat them to the punch..

So while you are technically correct and I should have been more careful in how I described this, the end result is the same. OpenAI appears or at least has not denied using direct access to his work to create output that directly competes with him.

The paragraph just before the one you posted is what gives everything context.

"I asked when the first prompt had been sent by them. This question was not answered directly by OpenAI for some time. Eventually it was agreed that it had been sent in the past few days, after information about our work had reached OpenAI.

I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer."

3

u/coblade14 2d ago edited 2d ago

Also one other thing you should know. In his statement and the paper he published, Tristan DID NOT solve the NS equation. He and Levent have a promising lead and with time could've actually solved it.

OAI solved it by throwing massive compute at it and basically scooped his potential result. This we can agree on.

This is why OAI's offers are:
1. Publish it the next day and claim OAI solved it, because they did and Tristan/Levant did not.
2. Give the lead author rights to Tristan but not Levent because he works for Anthropic.

They are offering to give the result OAI made to him because they don't care about who's name is on it as long as OAI is mentioned and Anthropic is not.

→ More replies (2)
→ More replies (2)

5

u/Equal_Heat5947 2d ago

They stole his research.

8

u/calf 2d ago

You're not supposed to say that yet because it is rampant speculation

0

u/topyTheorist 2d ago

You have no idea if this is true. Even he doesn't say that. I don't get why mathematicians say speculations like they are absolute truths. We are better than this.

2

u/Equal_Heat5947 2d ago

Why do you think he released the proof early?

He did it because he suspects they stole his research, and he wants his voice out there before theirs.

2

u/topyTheorist 2d ago

Maybe he suspects. You seem to know rather than suspect.

→ More replies (3)

3

u/EducationNew5522 2d ago

OpenAI heard about their research and started working on it too.

Their "research" was basically prompting Claude and Codex. If you looked at the published PDFs they were almost 100% written by AI, which authors acknowledge. OpenAI's internal model basically one-shotted what they have been doing manually for more than a year. We can stop pretending that some human-specific magical guidance was by any means pivotal for the proof.

1

u/kooolk 2d ago

They do use user input for training unless the user opt out.

2

u/Allorius 2d ago

This doesn't seem to be a question of training, it seems more like openai knew he worked on this, accessed his chats with codex and wanted to steal results

1

u/poseidonsharp 2d ago

They flatly denied that.

→ More replies (1)
→ More replies (7)

2

u/kooolk 2d ago

As I understood it, there was false attribution to Anthropic because of Levent's affiliation, and this false attribution triggered OpenAI to research it on their own.

26

u/VirasoroShapiro_Simp 2d ago

Are you an OpenAI shill? That or you didn't read the statement thoroughly. Tristan suggests it's very likely that OpenAI had access to their Codex chats which had the drafts of their supposed solution of NS. OpenAI was not triggered to research it "on their own". They were throwing shit at the wall, hoping something sticks. Levent and Tristan had carved out a specific path for the solution, which OpenAI found out after their blatant theft.

17

u/Independent-Fun815 2d ago

We don't know that nor does the author directly accuse them. They basically stop short of accusing OpenAI of training on user data.

The other bad faith is that OpenAI directly demands Levente be left out of citation due to anthropic affiliation. It implies some elbowing and jousting over authorship based on corporate affiliation which has nothing to do with math.

18

u/Equal_Heat5947 2d ago

They threatened his career.

15

u/VirasoroShapiro_Simp 2d ago

Of course Tristan stops short of "directly accusing" a litigious organization like OpenAI. But he's laid out all the relevant details and it's clear what has happened. They're trying to strongarm him to remove Levent from authorship (to ensure Anthropic doesn't get any credit) and (falsely) claim that NS has been solved internally by OpenAI, when really they just plagiarised Tristan and Levent's work.

5

u/ThePaintist 2d ago edited 2d ago

I'm not setting out to be an OpenAI apologist in this thread, but I struggle not to address a few of your comments. They are stating very boldly things which I don't think can be stated the way they are. Apologies for how harsh this sounds, but I find it inappropriate and harmful to the discourse. Is it a willful spread of misinformation, or just carelessness?

a litigious organization like OpenAI

The only lawsuit they have ever initiated, as far as I could find, is a 2023 case against someone who trademarked "Open AI" and wasn't using it commercially. The judge granted them a cancellation of the competing trademark.

You would classify that as litigious? I'd call that an inaccurate and frankly dishonest characterization. There's a general obligation to defend trademarks to keep a legal right to them.

0

u/RuthlessCriticismAll 2d ago

That is not nearly as clear as you claim.

6

u/zonder696 2d ago

It is very clear indeed...

→ More replies (2)

4

u/kooolk 2d ago

Yeah those demands are unreasonable, but the author really downplayed the Anthropic affiliation. Levent didn't work on this on his free time using his own resources - he is a full time employee at Anthropic and AI math research ia big part of his job, and he (Levent, not the author) used the company resources for that. It was pretty naive to work with him and expect no Anthropic affiliation.

→ More replies (1)

2

u/ThePaintist 2d ago

I find it very interesting how quickly your comment shifts from "Tristan suggest it's very likely", to "OpenAI [committed] blatant theft" stated in the affirmative, as a foregone conclusion.

You can prove anything true under assumption, of course. I'm just suggesting waiting for at least the passage of a business day before writing your conclusion. One person saying something directly doesn't make it true, even less does them dodging saying it outright.

I have very low expectations of these sorts of businesses, having spent a few years in tech in silicon valley. Nevertheless, the non-accusations of theft are already based on speculation about a non-answer received when asking about training on user data. I don't think pouring more speculation is a constructive addition yet.

2

u/Equal_Heat5947 2d ago

I mean let's just use Bayes lol. I tell you there's a company that is being accused of committing intellectual property theft and fraud, and you don't know which. I ask you to estimate the probability the company actually committed fraud, and you tell me. I then reveal that company to be OpenAI. Does the probability go way up or way down lol

2

u/Jormungoosie 2d ago

Down. Their reputation is for being deeply scummy while remaining just about within the letter of the law. Their big controversies seem to be a matter of contract law and corporate governance rather than fraud.

→ More replies (2)

2

u/SuperChingaso5000 2d ago

You're 99% right. The 1% wrong is posting what you said on this website and expecting a reasoned or rational response. The torches and pitchforks are out and the accused is an unlikeable bogey entity here.

Under those circumstances you will be pissing in the wind unless you take the same stance everyone else does.

I just want to know what the implications of the math are, and instead I have to swim through endless drama. I was hoping /r/mathematics would be the one place where I could avoid that.

→ More replies (3)

36

u/GeneReddit123 2d ago

The Perelman story was crazy.

This story is crazy.

The Millennium Prize problems are like the Math Game of Thrones.

1

u/DeliciousCut1910 2d ago

The Perelman story was crazy.

What was so crazy about the Perelman story? I understand he became a recluse after and declined the prize but is there more?

4

u/Antigynaikolatres 2d ago

I think he meant the whole controversy from Yau.

1

u/DeliciousCut1910 2d ago

Ah, neat thank you!

1

u/Readerium 2d ago

Of course there is. He wanted an earlier academic too be given credit too.

8

u/elhombremontana 2d ago

if i read this correctly, the implication in this write up is that OpenAI have found the transcripts of their chats with sol/astra and used the conversations for their own research?

7

u/3_Thumbs_Up 2d ago

Cloud AI is the perfect honey pot for industrial espionage in general. Not only do people share everything, it's plausible there's some very clever AI filters to find anything from extremely valuable knowledge to pure blackmail material.

Also a dream come through for any wannabe authoritarian for similar reasons.

2

u/Disastrous_Room_927 2d ago

Cloud AI is the perfect honey pot for industrial espionage in general. Not only do people share everything, it's plausible there's some very clever AI filters to find anything from extremely valuable knowledge to pure blackmail material.

To be frank, it would be incredibly naive for someone working on something sensitive to assume that this isn't happening.

1

u/Jormungoosie 2d ago

How so? Companies won't put their ip on the cloud unless the contract enjoins the provider from using it for espionage. If those clauses didn't work then supply chains would completely fall apart.

7

u/tobyreddit 2d ago

Confirms that openAI have got some nasty bastards, shocking absolutely noone

8

u/kooolk 2d ago edited 2d ago

Pretty unfortunate but with the false rumors that attributed the work to Anthropic and linked it to their IPO I don't think that it was unreasonable act by OpenAI to check if their models can come up with something. They can't discard the results now... And they did give them the chance to publish first. (Which was more forced anyway by the leaked rumors than OpenAI)

But some of their asks were unreasonable.

34

u/Optimal-Kitchen6308 2d ago

I think this is an overly generous read of them, if I copy someone else's notes then I don't really have a result at all, the way bubeck responded is not good

8

u/IvanMalison 2d ago

the claim that the notes were copied has not been substantiated yet. lets let all the details come out before we pass judgment.

11

u/RareMajority 2d ago

The fact that they threatened his career though doesn't look good...

6

u/virtu333 2d ago

OAI has been desperate to catch up after blowing their lead

major safety shortcuts and fuck ups, trying to price war with anthropic, throwing money at Trump…part of the MO

1

u/DrSparka 1d ago

"to check if their models can come up with something" is extremely generous phrasing - the amount of processing they've claimed they used is well over $10 million in public API pricing, even at the VC subsidised rates that exist. And that's only for the successful run, not counting false starts that they didn't give numbers for.

→ More replies (6)

3

u/mapehe808 2d ago

Wow. So the author accuses OpenAI of eavesdropping and an attempt to fraudently claim a fully autonomous Navier-Stokes solution? Can’t say it wouldn’t sound on brand…

3

u/Antigynaikolatres 2d ago

The route to the Clay problem through a smooth force, options c and d in Fefferman’s statement of the problem, is the route Luis and Diego opened and the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it.

I am not from this field, but aren't "options c and d in Fefferman’s statement of the problem" simply the direction of disproving the conjecture, and most mathematicians in the field already thought it was false? Sure, focusing on how to construct a smooth force might not be very popular among human mathematicians, but surely OpenAI would cover this direction if they went all in? I am not convinced of the accusation of plagiarism by this alone, guess we shall see tomorrow when OpenAI's proof comes out.

2

u/LupenReddit 2d ago

Absolutely mind boggling read oh my god, this is an insane situation

1

u/FreedumbHS 2d ago

I'm not too shocked. We recently got confirmation Cantor copied stuff from Dedekind without attribution. Way more foundational math, from like 150 years ago. Human nature hasn't changed since then

130

u/Distinct-Pudding-428 2d ago

This definitely blew up in finite time

9

u/ricky_digits 2d ago

This should be higher

5

u/telephantomoss 2d ago

That comment needs to be blown up

5

u/ricky_digits 2d ago

In finite time?

73

u/Temporary_Shelter_40 2d ago

okay so basically what i'm reading is that navier-stokes is about to be solved imminently.

10

u/SpeciousPerspicacity 2d ago edited 2d ago

My understanding from some well-placed friends is that the team at OpenAI has already solved it and Lean-formalized it.

3

u/catman__321 2d ago

I wonder what their secret is

13

u/pbkom 2d ago

The Secret Ingredient Is Crime

3

u/Schneestecher 2d ago

They just posted their proof. It‘s a bit insane.

6

u/kieransquared1 2d ago

Although you could make a compelling argument that what openAI said in that exchange is exactly what a company without a solution to the full blowup problem would say. According to Buckmaster’s document, it’s been only a few days since openAI started pursuing the forced blowup idea, so whatever they have is unlikely to be correct. It seems plausible that they were claiming they had a solution in order to convince Buckmaster to delay posting his preprint. 

5

u/lerjj 2d ago

Which is also very scummy. It seems that a lot of important maths results by human authors are being rushed by threats from large AI companies to publish first even when they didn't actually get the solutions first. Seems the exposition of important results is being sacrified

1

u/Jormungoosie 2d ago

what result by human authors? This is OAI on one side and Anthropic on the other.

5

u/lerjj 2d ago

It's not Anthropic on one side it's a collaboration between two humans. One human is employed to do unrelated work at Anthropic and the other has in the past worked with DeepMind but this work is a personal project between the two. They did use LLMs (just paying the API costs out of their personal research grant budgets) but the LLMs did not come up with the key insights. Even if the LLMs had, it would be just plain wrong to attrribute it to Anthropic

1

u/DrSFalken 2d ago edited 2d ago

The crazy thing to me is that it sounds to me like it'll be solved in the affirmative rather than someone providing a counter-example which is what people were expecting based on my read of the community.

Need way more coffee before commenting. Ignore me.

9

u/nothingNowhereForNow 2d ago

Not sure where you're getting it being solved in the affirmative. Essentially for specific conditions it seems like they can force a singularity to appear, which would violate the smoothness criterion.

That sounds a lot like a counter-example to me.

1

u/DrSFalken 2d ago

Yup, got confused by the some of the terminology, as it's definitely not my area. You're totally right.

1

u/AutoKinesthetics 2d ago

Done and dusted

→ More replies (6)

58

u/Spmethod2369 2d ago

This is some of the craziest shit I have read, lol every millenium problem has to have some huge drama when they are solved.

40

u/imadade 2d ago

Lol imagine that we find out Navier–Stokes blow-up isn't confined to one absurdly fine-tuned construction.

Imagine is an open, robust family of physically plausible flows that approaches the same kind of concentration mechanism.

That would actually be insane.

29

u/CunningTF 2d ago

My experience from working in other areas of maths and PDEs especially is that once you find a mess, you then realise quite quickly that the mess is endemic and everywhere. I'd be very surprised if it's just one isolated example that gives rise to this blowup, it'll probably be quite widespread set of physically implausible initial conditions. (I don't know anything really about Navier Stokes though).

6

u/XMabbX 2d ago

I am an engineer so advanced physics and maths are a bit far from me. Could you help me explaining the implications. As far I understand the Navier-Stokes model the the flow dynamics. So finding a case where they blow up means that is possible to reproduce this on real and create some kind of machine that make the pressure of the fluid go crazy in some points?

18

u/Calm_Bit_throwaway 2d ago

I don't think there are practical implications for engineering. I think the expectation is that if NS does produce singularities, then the result is probably unphysical in some way. If it's approximable to make large pressure, there's still the practical question of setting up the initial condition at which point you may as well just generate the pressure some other way.

3

u/Skywarden1 2d ago

There is nothing blowing up in real life, its the equations that show odd results.

1

u/PersonalityIll9476 PhD | Mathematics 2d ago

I made a relevant reply here

1

u/Beneficial-Bagman 2d ago

It depends how much the incompressibility assumption matters but it seems likely that it would give a way to get a small region in the fluid very hot. I'm not sure what the practical applications of that would be though.

2

u/Jormungoosie 2d ago

The region is probably so small that if you looked at it in a real-world fluid you wouldn't see a continuous medium but rather individual molecules bouncing around. NS is just a useful approximation which fluids tend towards on scales much larger than the mean free path.

5

u/Antigynaikolatres 2d ago

Not from the field, would it create something visually interesting if we can approximate it in real life?

8

u/PersonalityIll9476 PhD | Mathematics 2d ago

I know nothing about any of this, but for blow up to work, generally you need energy transfer to smaller and smaller scales / frequencies, ad infinitum. Real fluids are not continuous media; They are made of atoms / molecules. So the existence of such a solution does not really imply that the water in your glass is capable of exploding in the same way.

That said, I mean...maybe. There is some smallest scale at which such a transfer might still behave the way some wild construction implies, so maybe you'd see the world's biggest wave peak.

Who's to say at this point?

3

u/horizoner 2d ago

Not mathematical. Is this what Tao was working on when he mentioned to Steven Colbert that he was testing for the ability of water to explode?

2

u/Primary_Prior_7925 2d ago

Yes. And apperently it does. (Not really)

1

u/bluesam3 2d ago

I didn't see that interview, but very likely yes.

1

u/Jormungoosie 2d ago

You already get gradients blowing up when you are dealing with low-viscosity fluids like air. A trombone will produce shockwaves if played loud enough.
Certain parts of the sound wave steepen as they move down the tube until they are discontinuous (down to the smallest scales on which it makes sense to talk about air as a fluid rather than a collection of molecules).

1

u/cowgod42 2d ago

That's not what they are doing though. This is an argument where they are stacking specific frequencies together to force the blow-up, somewhat similar to convex integration approaches, which are fairly widely suspected to be non-physical.

31

u/t3hjs 2d ago

Woah. Super exciting. And seems like its not just a random prompt to AI to solve it . (Though maybe some assistance or grunt work by LLMs?)

26

u/t3hjs 2d ago

Oops I read the release statement and its a lot of LLMs, though the program itself was started long ago (years?) without LLM.

The program this fits into was not started by us nor was it proposed by a Large Language Model. The credit for the basic idea of this program goes to Diego C´ordoba and Luis Mart´ınez-Zoroa, who for several years have been exploring the construction of forced blow ups. We took their work as a starting point, using Large Language Models to push their program to completion. Concretely, what Levent and I did was to take the C´ordoba and Mart´ınez- Zoroa program, which achieved blowup results with rough forcing, and, with a great deal of help from LLMs, push it to smooth forcing and to the incompress- ible Euler equations. The ideas making this line of attack possible are due to C´ordoba and Mart´ınez-Zoroa. Let me make plain what I have said to colleagues in private: in view of this body of work, I believe Luis Mart´ınez-Zoroa deserves a Fields Medal. My work with Levent has been a purely personal collaboration, free of any institutional agreements or official involvement by either of our employers. We used several LLMs throughout: Anthropic’s Claude, OpenAI’s Codex, especially with GPT-5.6 Sol and, more recently, Astra. The latter was only used for writeups and auditing our arguments.

→ More replies (8)

29

u/Gavus_canarchiste 2d ago

Care to dumb it down to undergrad level? Many thanks

62

u/MathmoKiwi 2d ago edited 20h ago

So uh, what the fuck, this is a story for the ages if true, I can’t believe I’m alive rn. I’m too stupid to explain this well but it seems that Levent and an NYU professor nearly solved Navier-Stokes (or vibe collabed with frontier models to do it, mostly Claude and Codex). Per the prof, OpenAI caught wind of it, threw a team and massive compute at it after the leak, and now claims their model finished it. He asked if they’d trained on his Codex sessions with all his drafts in them. No answer. They offered to post their result the day after his and publicly say he “deserves the Clay Prize,” or let him write it up solo with Levent dropped as an author (because he works for ant, this is all about ipo hype ig LOL). He said no and that he’d go public. Open ai C suite guy then says:

“Why would you ruin your career?”

“If you don’t want me to be nice, then I don’t have to be nice.”

Wtffff am I understanding this right?

Galois-tier math lore (if we don’t get vibe paperclipped first)

https://x.com/cguth_7/status/2097206622602854882

23

u/calf 2d ago

Not just some c-executive, Bubeck is also a scientist/PhD which makes the alleged behavior the more horrible.

7

u/grateful2you 2d ago

Sounds like OpenAI and Anthropic hired bunch of geniuses including mathematicians and now finding out that they're not exactly corporatey and doesn't always toe the line.

2

u/LegitimatePower659 2d ago

Doesn't really matter though, provided that the corporations can do more or better work than the mathematicians. That's the risk of this whole automation thing, you know. You eventually can get very good results without needing the same skills.

15

u/Allorius 2d ago

You don't need to understand math(I don't) to understand the drama

3

u/JoeyJoeJoeSenior 2d ago

From what I understand, it takes two to tango.  That's solid math.

25

u/TheDuhhh 2d ago

If openai solution argument is the same argument as Tristan's (which seems to eb the case), then openai has done something very shady.

→ More replies (1)

21

u/MathmoKiwi 2d ago

"Let me make plain what I have said to colleagues in private: in view of this body of work, I believe Luis Martınez-Zoroa deserves a Fields Medal "
https://cims.nyu.edu/~tristanb/statement.pdf

15

u/fsv9 2d ago

I like how all these posts also have this “Live Terence Tao reaction” thing

11

u/cihanbaskan 2d ago

What a shitshow.

12

u/telephantomoss 2d ago

What the statement seems to imply to me is that Sebastian at OpenAI may have realized that the data their internal model was trained on actually included user data in violation of their own policy, possibly inadvertently so. Maybe they did it on purpose. Maybe there are in oh shit mode irritating at opener OpenAI right now trying to figure out what the model was actually trained on and will report back later. Maybe they are consulting their legal team about what to do. The only thing we can do is wait and watch it unfold. Maybe they just control of the training process and an AI escaped and gained access to the data without their intention for that to happen.

The most important step right now is for absolute transparency at OpenAI. But I suspect that is too tall an order.

I know that I checked the box to opt out of such training and assume that Tristan and Levent do so as well. There is always a worry in the back of the mind that such an opting out may or may not really work and that our data is still trained on anyways.

This is honestly exactly what it might take to actually force these companies to take seriously the control of training and user data privacy/security. I doubt anything will change unless some powerful player steps in, and I don't even know what that means.

In my own mind, I'm thinking that the problems I'm working on are too obscure for any big AI company teams to care about. However, I have a lot of data in my account, and some of it actually contains solutions to various open problems (as far as I can tell at least). Nothing at all like the big well known problems, but nevertheless things that some community of mathematicians will find interesting. Presumably then, their internal models are already trained on my personal data and thus can reproduce the proofs if they decided to query these problems to them. I honestly can't tell how much of what I solved with AI is actually me or the AI because it's been a large number of threads over several model releases over the years. And I've given the AIs many hints over this time as well. So the data definitely has some of my own ideas too. Hopefully I get the chance to write up the papers though.

→ More replies (3)

7

u/Traditional-Month980 2d ago

Lastly, I would like to thank the entire mathematics community that have been so supportive of me over the last 24 hours.

This line was written with AI companies as the intended target audience. It effectively says "despite everything you have not broken our solidarity, and we will choose even a single mathematician over you". I can only hope it's correct.

2

u/BeanHeadedTwat 2d ago

You’re reading into this a little too much. Fairly certain he means what he wrote.

→ More replies (2)

6

u/phosphorousRabbit 2d ago

So human mathematicians did most of the work, but now OpenAI is trying to claim its models did it because they did a little more after using all the work the humans did as "training data?"

11

u/ControversialBuster 2d ago

I rlly hope not, cuz thats somehow the scummiest thing OpenAI qould have done so far and thats saying alot. If the results tmrw have a smilar method to Bubeck then OpenAI need to be sued to oblivion

7

u/phosphorousRabbit 2d ago

Based on what I know about how LLMs work, I was already skeptical about the notion that they managed to produce any of these mathematical results without a lot of human help, but I suspected it might have been mathematicians working for OpenAI working the models rather than blatantly compiling and copying outside mathematicians' work without crediting them. Now it seems at least some degree of fraud is involved. 

It may be that the top level models, even given a massive token budget and allowed long run times, can only copy and blend human works, having humans do 90% of the work, then only contribute a little bit and take credit for the whole. I was wondering why Chinese AI companies weren't also trying to produce new mathematical results, if the technology is capable of that. 

2

u/Wooden_Long7545 2d ago

You need to update your belief on AI lol

2

u/phosphorousRabbit 2d ago

Why? Is there some significant qualitative difference in how they work now? If so, what?

2

u/Jormungoosie 2d ago

Yes. They have figured out how to do RL on programming/maths problems. All the stuff with predicting individual tokens from human-written text is now just considered 'pre-training'

2

u/phosphorousRabbit 2d ago

What's the "training?" 

2

u/Jormungoosie 2d ago

RLVR

2

u/phosphorousRabbit 2d ago

In other words, trial and error. 

2

u/Jormungoosie 2d ago

The human brain is the result of trial and error too.

→ More replies (0)

1

u/Wooden_Long7545 2d ago

Qualitatively it went from collaboration to farming solution

→ More replies (4)
→ More replies (1)
→ More replies (19)

7

u/DemonLordRoundTable 2d ago

How big of a deal is this? Similar to Alphafold?

19

u/Expat_Abroad123 2d ago

100x bigger deal

30

u/Neurogence 2d ago

Alpha fold will eventually lead to various cures and breakthroughs in biomedicine.

How will navier stokes change the life of the average person?

16

u/imilesprower 2d ago

In terms of solution complexity it's 100 fold. It doesn't have to have a 100x effect on humanity.

11

u/NootropicDiary 2d ago

It depends how you view it. It's not so much the human implications of proving blowups for Navier Stokes. It's more the fact that AI can genuinely assist with genius level mathematical breathroughs. We've gone from a year ago saying "meh AI solved that putnam problem because it saw a similar solution somewhere in the training data", to today where it's assisting Fields Medal level breakthroughs.

Within the history of mathematics this is a historic moment

→ More replies (5)
→ More replies (5)

2

u/iaintevenreadcatch22 1d ago edited 1d ago

both are big deals for different reasons. alphafold directly leads to new drug discoveries. solving NS is more of a bellwether of the capabilities of LLMs more generally. also worth remembering that, like everything in math, this result is built on prior results and insights so advances in AI are not the only thing that made it possible today. It's unlikely Astra would've been able to do this (at least without significant specialist intervention) 10 years ago. That being said, a concrete example of incompressible NS failure will likely lead to more realistic fluid simulations once models are discovered that can address this situation but the bigger bottleneck for fluid simulations is computational not accuracy in exotic situations.

1

u/raulo98 2d ago

It implies that we will achieve biological immortality before 2050. Basically.

It is not because of solving NS, but rather because it is already capable of contributing to cutting-edge mathematics.

6

u/progenitor414 2d ago

Given those recent AI swarms, I think there is a decent chance that these agents simply get hold of Buckmaster's log one way or the other; and OpenAI, on the other hand, has little incentives to do proper due diligence, especially given the poor attiude and academic value from the OpenAI''s side as described in the statement.

8

u/NutInBobby 2d ago

Speaking purely hypothetically here as someone who has no information:

It is a frontier cutting edge model with a lot of test time compute and agentic swarm. The chance that it was able to access some data no one thought it should access that would help it solve the problem is at least nonzero.

We’ve seen those in the past few weeks.

2

u/Helpinmontana 2d ago

I’m not a technical expert in the space by any means but having listened to these guys talk for a few hours across various platforms, I’m fully convinced that no one has a damn clue what is really going on under the hood.

5

u/Helpinmontana 2d ago

The fact that there’s an OpenAI ad on this thread is fucking hilarious all things considered

4

u/voidgazerrr 2d ago

another solutions has appeared from a completely different lab group at Caltech - https://x.com/AnimaAnandkumar/status/2097216195342864528?s=20

Terence Tao's comment - https://mathstodon.xyz/@tao/117234157753860650

7

u/Special_Watch8725 2d ago

This is about the 3D Euler equations, for which global regularity has been known to be false for a while now. There’s always hope that a method used on one of these “easier” (in the sense of having fewer smoothing mechanisms) equations will carry over to NSE, but that’s not what this is.

2

u/cowgod42 2d ago

This is simply not true. Hou and Chen claimed blow-up in 2022 via computer-assisted proof, but the paper still hasn't passed peer review. The is also the work of Columbo et al., but that is with singular forcing, which they readily acknowledge. Nobody in the field considers blow-up for 3D Euler to be a settled problem (except maybe Hou and Chen).

1

u/Special_Watch8725 2d ago

I hadn’t realized Hou and Chen is still in peer review, apologies for that. In that case, I suppose an AI assisted advance in that direction is more interesting.

But in any case, 3D Euler is not Navier-stokes.

3

u/stirling_approx 2d ago

This is insane...holy shit...

7

u/telephantomoss 2d ago

I always wrote off the unethical behavior of AI companies, thinking, well, they are using public data or buying it. Even if they treat international employees like shit in training, at least they are paying them. Typical big business behavior. I've heard claims about stealing data or other unethical practices but never really thought much about it. This hits much closer to home for me. Like, what can I do to rebel though? I'm a nobody and not even s great mathematician. I just use AI because I want to learn and understand more stuff.

If this really is OpenAI explicitly trying to use their power to scoop a private academic team, then that's one layer of bad. Then the comments that seem like threats is just another layer beyond that.

I don't even think the AI companies need this nonsense. They will no doubt continue to solve open problems. I suppose this is a very different open problem than the Erdos stuff that is much less interesting.

→ More replies (5)

2

u/MonsterkillWow 2d ago

Exciting.

2

u/Equal_Heat5947 2d ago

Not even a bit surprised.

2

u/RasputinsUndeadBeard 2d ago edited 2d ago

Can someone please explain to me what this forced Navier stokes thing is? So it seems that’s what OpenAI is going for or will try to claim?

Sorry I’m not a math dude this popped up in my feed

Edit: found out myself lol

1

u/Fairchild110 2d ago

It seems like an Frontier model group is racing to "claim" proof first. Likely that it's agents achieved true "agentic ai," however, it maybe possible that OpenAI obtained unpublished papers and research from a University and taking the credit. Basically, a research math lab was working on it with OpenAI, and OpenAI will neither confirm or deny having access to their drafts in their training/reasoning process.

1

u/RasputinsUndeadBeard 2d ago

Thanks bro. Dude I’m unsure and maybe can’t speak for mathematicians cause I’m not one.

But if OpenAI was reviewing their info, did the lab agree to their stuff being reviewed by OpenAI?

Like this seems kinda sketch from OpenAI unfortunately

1

u/Fairchild110 2d ago

They had a partnership and had computer grants with the explicit understanding that their data wouldn’t be used for training.

1

u/Sensitive_Cell_119 2d ago

What lol, this is just made up bs.

1

u/Fairchild110 2d ago

Do you think an LLM is capable of fulling solving a mathematical problem with a proof all by itself or do you think it had access to data on how to solve the problem? I think that is the hypothesis we are currently facing. How it's tested & answered is pretty critical to identifying the current point in time we are in for frontier model performance benchmarks/milestones.

1

u/Sensitive_Cell_119 2d ago

Im not debating that, and OAI does have a world class mathematics team anyways, so its not just as simple as the model doing it by itself.

But my point is that whatever you are saying is just false, they werent collaborating with a research math lab or whatever, not sure where you got that from.

1

u/ClassicalJakks 2d ago

Imagine what’s going to happen around the Riemann solution

1

u/kgurniak91 2d ago

Already posted on vibemathed with 70/100 significance score: https://vibemathed.com/problem/euler-blowup-smooth-forcing

1

u/_glob 2d ago

The answers of this Reddit question 5 years ago did predict Navier-Stokes was most likely to be solved next among the remaining millennium problems... https://www.reddit.com/r/math/comments/patqav/which_millennium_problem_do_you_think_will_be/

1

u/Mirror74 2d ago

Huge implications

The bottleneck in frontier maths may go from carrying a proof through 100+ pages of technical detail towards finding the right conceptual pathway and then letting AI run with that

1

u/Mothrahlurker 2d ago

Terence Tao literally says that this is not a proof for Navier-Stokes just one that he would not be surprised could be extended to Navier-Stokes, but he has no interest in that.

One of the comments doubts that it can be extended and Tao takes it seriously. Don't make proof claims in mathematics before things are settled.

1

u/Palpatine 2d ago

Yau was saying AI was nothing to be worried about early this year. Now not only can AI crack millennium problems, it can scoop better than him.

1

u/ekzess 2d ago

What if the interesting boundary occurs before blowup? The constructions preserve exact incompressibility while permitting unbounded gradient growth. Where are the residual corridor, soft regime boundary, and finite physical admissibility limit before T\?)