Rendered at 20:47:35 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
n2d4 2 days ago [-]
Drama/accusation summary:
- Aug 15th: Tristan Buckmaster & Levent Alpöge make progress on a few important math problems, "finite-time blowup with smooth forcing for incompressible porous media, for Boussinesq, and for 3d incompressible Euler."
- they do NOT have a proof for the $1,000,000 Millenium Prize problem. BUT, they do claim to have a proof for a similar (non-Millenium) Navier Stokes problem that could help lead the way there
- Levent works at Anthropic, but this research was independent of his work there, with a mix of GPT and Claude models. Tristan is not related to Anthropic.
- Early Sep: Rumor spreads to OpenAI that Anthropic solved a major problem. Tristan emails OpenAI to clarify, without revealing the problem they solved or how they did it.
- After hearing of the rumor, OpenAI started researching Navier Stokes with a new internal model.
- Sep 6th: OpenAI's Sebastien Bubeck tells Tristan that they solved the $1,000,000 Millenium Prize Navier Stokes problem. The approach is very similar to Tristan & Levent's approach to the non-Millenium problem.
- Tristan is suspicious of the timing, as only few others were trying this approach. OpenAI says the model didn't access his user data directly, but leaves unanswered whether Tristan's chat conversations were part of the training.
- OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic.
- Sep 8th: Tristan refuses to remove Levent, and rushes to publish their results independently.
sigbottle 1 days ago [-]
It's specifically the last two bullet poitns
- Tristan is suspicious of the timing, as only few others were trying this approach. OpenAI says the model didn't access his user data directly, but leaves unanswered whether Tristan's chat conversations were part of the training.
- OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic.
These two bullet points are extremely suspicious if you were honest. Like I'd imagine for OpenAI, they'd love to pump their chest and not even give Tristan credit - "no, we did it, GG mathematicians". It's this weird hedging half-assed measure, especially with the desire to remove Levent, that makes it suspicious.
hkmaxpro 1 days ago [-]
Both Sam Altman and Sebastien Bubeck admitted they only want Buckmaster to be the lead author on a rewrite of the OpenAI proof.
A wake up call for using OpenAI models. If you discover something with their model and you work for a competitor, they “felt it would be inappropriate” for you “to author OpenAI’s work”.
esalman 15 hours ago [-]
I work in catastrophe risk modeling and it's a multi billion dollar industry.
We often chat where the business might be heading in future. An uncomfortable scenario is what if a frontier tech company decides to offer our customers the same products that we do.
There's a lot of pressure on AI adoption so the company has partnered with various tech companies to build intelligent systems on top of proprietary data and mathematical models.
If OpenAI is indeed using customer data to train their models to win a $1m prize, then it throws a giant IP question at the partnerships that affects multi billion dollar businesses.
DeepSeaTortoise 14 hours ago [-]
> If OpenAI is indeed using customer data to train their models to win a $1m prize
Is that even a question? Of course everything not kept on premise at gunpoint is going to be trained on. The chances of getting caught are 0 and the consequences of getting caught are 0 (as we've seen with copyright laws going from sending people to jail for years to unenforced within months). Yet the benefits are through the roof. Your customers aren't going to pay for having the very same data vibe enriched twice, it's exclusive, extremely high value data your competitors will never have access to.
rickdeckard 13 hours ago [-]
Agree, I think the practice is also very clear from the overall strategy of AI-companies and their ToS:
Scale with subsidized pricing as fast as possible to gain more user-data for training --> Own the better model --> scale pricing.
Scanning social media (e.g. Twitter, Reddit) posts only give a glimpse into the thought-process, chat logs on-scale give you the actual process in machine-readable format.
There's a reason why Google considers the Emails of Spirit Airlines to be worth millions of dollars [0], they give insights into a process, not just into the results...
> - Tristan is suspicious of the timing, as only few others were trying this approach. OpenAI says the model didn't access his user data directly, but leaves unanswered whether Tristan's chat conversations were part of the training.
The question, for AI customers, is when they build products using services of AI-companies, would AI-companies engage in theft of customer data for use in training?
marcosdumay 3 hours ago [-]
If you still had that question, you can answer it now.
But honestly... "Will the company that was entirely built over illegally acquiring data use some data that is legal to use and is right on their front, or will they not do everything they reserve the right to do?" is a really bad question for one to even ask.
rickdeckard 7 hours ago [-]
The fantastic grey area that was engineered over the past decade is "profiling", so my guess is the answer will be "we didn't use your customer data for training, but we cannot rule out that it has been used to create profiles of your customers to train our model"
AlexCoventry 12 hours ago [-]
I'm not a lawyer, but I think the comparison to copyright law is invalid. Such a dispute would be governed by contract law.
barrkel 12 hours ago [-]
This would be contract law, and it would also be a huge reputational risk. All it would take is a whistleblower and there would be billions lost.
DeepSeaTortoise 11 hours ago [-]
Sure, but I highly doubt that there would be many people involved. And those who are, are probably quite interested in keeping it that way and not at all in becoming whistleblowers themselves.
You wouldn't want to decide what's worth training on and what isn't manually, so there is almost certainly an automated pipeline to do so (certainly at least for the free accounts and those that dont opt out of training).
Then there's the question if this pipeline only sorts through the data or also transforms it and to what degree. E.g. for removing personal details, locations, medical information and so on. The data that comes out of this pipeline might have VERY little information left in it a human could connect to the original input. Even worse, since we're talking about companies specializing in sota statistics, the input data could have been transformed into a representation that is very well suited to represent all the novel and interesting parts, but is awful at modelling all the things that could end up identifying where the data comes from (or causes legal liabilities otherwise).
In the end the only thing a potential whistleblower might even have a chance at observing in the first place, is whether a company's data enters such a pipeline or not. And I have my suspicions that the major AI companies operate at a scale and level of automation, that absolutely nobody has a chance at figuring out where anyone's data is at any point in time and what any specific piece of equipment is currently busy with.
So the only place to figure out whether data is trained on that shouldn't be trained on is by looking at whatever configurates every single system that could take a peek at some customer's data or the systems themselves while processing the data.
The latter would be such a huge violation of a customer's rights, no whistleblower is going to attempt that or admit to doing it.
And the configuration for the former could live just about anywhere, from regular config files to the CI/CD pipeline, pre-compiled libraries, kernel modules, modified vendor firmware, the compiler itself ... and probably plenty other scenarios you'd have to train an LLM on the ramblings of a crackhead to come up with.
So I'd say a whistleblower is pretty out of luck even becoming one.
visarga 5 hours ago [-]
You can just spin up deep research agents that ingest many sources at once to produce reports that don't replicate any one source too much. Since agents compare against sources they provide across-source analysis - what is the distribution of positions on this topic, is it debated or settled. Not truth, just summarizing, but I think this would be very useful for training.
Besides reporting on search sources you can also run the same queries on multiple LLMs closed book mode, and judge their distribution as well. It helps a lot if models are more aware of their knowledge holes. Scale it up for billions of topics if you have the pockets, the DR data is copyright free.
9 hours ago [-]
lou1306 12 hours ago [-]
Ok there is a non-zero chance that they could face a lawsuit and get fined for billions, but that chance is not 1 either: there is always a chance they get away with it. And even if they don't, if in the meantime they farm 10- to 100-fold that amount of money by just breaking the law, it's still a no-brainer for them.
desterothx 11 hours ago [-]
Billions lost, while waiting for their trillion ipo. Im sure they would manage...
theturtletalks 12 hours ago [-]
When these LLM companies were pirating content to train and it wasn’t punished at all, I knew the rules don’t apply to them.
But don’t worry bud, instead of the authorities going after actual corporations admitting to actual crimes, we’ll just ban CloudFlare IP addresses for everyone during La Liga games to battle piracy.
noir_lord 10 hours ago [-]
And require real ID to do almost anything on the internet "unintentionally" enriching their data sets by tying what you asked/where working on to you specifically as a person.
chinathrow 12 hours ago [-]
> The chances of getting caught are 0
I'd say non-zero, as seen in the current state of affairs.
bryanrasmussen 10 hours ago [-]
sort of agree, but also sort of think an accusation with lots of people arguing is not exactly the same as being caught.
12 hours ago [-]
steve1977 10 hours ago [-]
Pretty much everything or at least a lot of what you use as an OpenAI (or Anthropic or whatever) customer was once someone elses product that just got appropriated by OpenAI.
Why should your business be any different?
motbus3 10 hours ago [-]
> We often chat where the business might be heading in future. An uncomfortable scenario is what if a frontier tech company decides to offer our customers the same products that we do.
I feel this is exactly what will happen as they cause all sites to go closed source to protect their intellectual property and the AI companies offer only biased information. They are replace the business on internet model by bankrupting everyone with their own tools.
This is predatory pricing under most antitrust laws (imho, not a lawyer)
and it is very easy to do when you dont need to pay for the raw material.
KoolKat23 6 hours ago [-]
This is the business case already. And has been the case with tech companies for a long time. Your phones built in photo manager replaced a lot of what Photoshop does.
noir_lord 10 hours ago [-]
> If OpenAI is indeed using customer data to train their models to win a $1m prize, then it throws a giant IP question at the partnerships that affects multi billion dollar businesses.
I mean how could you expect them to not given they've trained the existing models on effectively the sum total of all human knowledge available on the internet without regard to copyright/ownership of that material.
It's a little trite but this absolutely runs into the "Frog and the Scorpion", it is simply in their nature.
KoolKat23 6 hours ago [-]
I mean this is a basic question, is your data used to improve models, did you opt in or not. There's nothing crazy here and it's not identifiable. You'll just conveniently find the next model iteration knows how to do it.
Any enterprise worth their salt already considers this stuff.
Catloafdev 23 hours ago [-]
These responses seem to me to make it abundantly clear who's telling the truth here. I wonder who this fools.
It would be extraordinarily easy to simply say, this model was not trained on your work, if that were the case.
It's telling that they refuse to acknowledge the root issue here, and are attempting to shift the conversation elsewhere.
rlt 15 hours ago [-]
"which is when I said that I did not understand why one would risk their career [over unfounded accusations]. Genuinely, at that moment, I was trying to care for him"
"Our aim was to see whether our system was also capable of this impressive feat"
"OpenAI's intention was to do everything possible to celebrate their mathematical achievements and the heroic efforts that they made on Euler"
For some reason I have a hard time believing people when they use language like this.
rickdeckard 13 hours ago [-]
phrases along the lines of "I don't want you to take harm while trying to accuse us" is quite an "impressive feat".
Maybe shows how fast these companies have grown without maturing. I can imagine old-world Intel and Microsoft acting in that way, but they were mature enough to not write it down like this.
However, Intel and Microsoft have been grilled in court for those practices and faced harsh consequences. I have yet to see this actually happening to any of these new AI-companies...
acchow 21 hours ago [-]
> It would be extraordinarily easy to simply say, this model was not trained on your work, if that were the case.
The Huggingface Attack revealed that making blanket statements like this is difficult and requires quite a bit of manual labor:
1) the agents spin for days and produce too much output to review
2) using LLMs to process that output skips many important details
Ergo, the agent could likely decide it would like to look through actual user data, hack its way into that data, and produce way too much output for a human to decide whether or not this occurred.
abofh 19 hours ago [-]
> The Huggingface Attack revealed that making blanket statements like this is difficult and requires quite a bit of manual labor
It requires humans to verify what agents have done.
Weird
nikisweeting 4 hours ago [-]
which is rapidly becoming a game of steganography cat and mouse
blini-kot 18 hours ago [-]
> hack its way into the user data
silly LLM, so ruthless in its pursuit that it puts real pressure on the innocent and the most open company on the planet
JoshTriplett 19 hours ago [-]
> These responses seem to me to make it abundantly clear who's telling the truth here. I wonder who this fools.
Anyone who just reads headlines, if the lie gets around to more headlines than the truth does.
vlovich123 22 hours ago [-]
I'm not sure it's so easy to tell whether a given piece of data was in a training run at their scale. It's entirely possible they think the answer is no, but on the off-chance that it could be, they'd rather not say no and then later it turns out they did and then they're claimed to be lying. If you were them, unless you could 100% rule it out, you'd hedge and say you can't.
Catloafdev 21 hours ago [-]
It may not be easy, quick, or simple to figure that out - absolutely fair.
But it is knowable. Their entire business is built around training models - they have the ability to know exactly what was in any given training run.
I guess time will tell.
m00x 21 hours ago [-]
It would be very difficult to say. It confirms that Tristan's data is likely part of the data the models use, but a lot of filtering, pruning, and transform goes into training.
Data has to be determined to be signal and not just noice, then it could go through processes of generating questions/answers from that data, then it RLHF's over this.
OpenAI have petabytes of data, all anonymized. It could take months to say for sure it was part of the training, and even more time to determine if it made any difference.
Catloafdev 21 hours ago [-]
Frankly, I don't buy this difficulty argument.
They know which model was used to come up with that particular idea.
A text search over the corpus of user data used in the training set can only take so long.
benmathes 18 hours ago [-]
I worked in the tracing and tracking all the thousands of data sets that got tweaked and permuted and changed hands between thousands of researchers and data engineers at a major lab. The data that goes into training runs is permuted so much from the OG data that tracing the lineage is not trivial (dramatic understatement).
And the difficulty is harder than just the extreme scale of text searching. but also explodes with organizational difficulty since there are so many people tweaking/shifting data independently upstream of the actual training run, and no they will not all add the telemetry you wish they did.
In the ideal, should it be this hard? Well, no, but that's org wrangling for you.
ashkankiani 12 hours ago [-]
It feels convenient to not spend time on engineering around tooling that could be used to answer a question like “did you violate copyright by training on X?”
benmathes 59 minutes ago [-]
Don't attribute to malice what is better explained by coordination headwinds in extremely large companies.
The engineering around tooling wasn't remotely the issue. It's getting all the (thousands?) data researchers mostly iterating on fine tuning datasets that would get bristly if they couldn't work outside version control in a python notebook iteratively tweaking their dataset that processed and reprocessed a few datasets until a threshold was reached.
The only _guaranteed_ chains of custody are down at the compute job and file read level. Which in a massively distributed computing job is... [redacted] nodes reading [redacted] fanouts of "datasets" that is just an abstraction over [redacted] individual files.
There's no malice here. Just way way way more complex than you'd first think.
10 hours ago [-]
Phemist 10 hours ago [-]
It definitely is solvable though. Data versioning is a thing and it can work quite transparently to the mutations done on the data.
To not know who made and who approved a set of mutations on data can easily become equally as mind-blowingly stupid as not knowing who made mutations to code. Code is a subset of data after all and search over (provenance of) data can be implemented as DAG traversal.
Not tracking data changesets like code changesets is certainly a choice, not really a constraint anymore. A similar choice I feel is implied by "extreme scale of text searching".
> no they will not all add the telemetry you wish they did
...is just a failure of the corporate policy surrounding data handling. Is git-for-data already considered telemetry?
Of course the truth is provenance of data is something best institutionally forgotten as quickly as possible. The only thing that matters is it's there, that the data has no history, and that's why it can be used in whatever way deemed necessary.
benmathes 1 hours ago [-]
All of your points are valid, and believe me I was trying to make them. The problem is one of culture. Most of the people doing this kind of work didn't like version control, and their work was really just running notebooks (like iPython or Google Colab) until a number was good enough and they'd submit the file for inclusion into training runs.
You can call it a policy failure, but these people were in very high talent demand and so top down dictates would risk "X people leaving lab Y for lab Z" headlines and morale hits.
I am not saying this is good. I am telling you that on the ground it is so much messier than it should be.
SantalBlush 6 hours ago [-]
This sounds like data laundering in effect.
benmathes 57 minutes ago [-]
The appearance of heroic efforts to get authoritative lists of what datasets went into which major model versions prevented actual data laundering (up to intent and mistakes). But don't attribute malice to that which is far far easier to explain with coordination headwinds: https://komoroske.com/slime-mold/
vlovich123 20 hours ago [-]
I think you may be underestimating how difficult a text search over their data is. They may have to build new mechanisms to do this. And what you really want is also an attribution of how much of a contribution a given corpus made which is a much harder question to answer; a single appearance of a chat probably has very little impact on the inference performance at this time unless it’s been explicitly preferenced somehow
Catloafdev 20 hours ago [-]
I don't think anyone really cares about 'the measured impact the data had on the exact result' - a question which is fundamentally difficult to answer accurately in the first place - but rather whether the data was used in training at all - which as Tristan described, was extensive, beyond simply a 'single chat.'
Can you explain the difficulty in engineering a search apparatus over a corpus of text data? Actually searching through it may not be easy, sure, but it's work that's doable, and creating an index is relatively trivial.
JoshTriplett 19 hours ago [-]
> Can you explain the difficulty in engineering a search apparatus over a corpus of text data?
My guess: "If we ever imply that's possible, people might start asking questions about all the other work we've ripped off, so the official answer is that it's impossible".
dd8601fn 18 hours ago [-]
If they literally can’t audit training data for a given model, they shouldn’t be operating.
JumpCrisscross 17 hours ago [-]
> If they literally can’t audit training data for a given model, they shouldn’t be operating
They can operate. They shouldn’t be claiming credit for discovering anything.
mapontosevenths 17 hours ago [-]
> A text search over the corpus of user data used in the training set can only take so long.
Did you notice the line in the article that says the models had access to an offline copy of THE INTERNET. Like all of it.
ummonk 16 hours ago [-]
Especially because the data that gets fed into training is first anonymized, so they’d need to look for navier stokes related stuff in the anonymized training set and then get make some sort of ad hoc process (with Tristan’s permission and sign off from legal) to compare the training data against his chats / Codex sessions to check if anything matches up. And that assumes his chats / sessions are still there, and not deleted to compare against.
varjag 11 hours ago [-]
If the method is indeed found in the training set it's not particularly important to de-anonymize it. You have the proof you need.
scott_weber 21 hours ago [-]
It should be quite easy: if they don't leak the user session data publicly, and don't commingle it with training data internally, how could it possibly end up in the training data?
What surprises me is they're not more boldly/plainly lying about it.
remus 15 hours ago [-]
How would they know for sure that some details were not part of some other training data they use? The authors may have discussed some tangential details on a forum for example, in which case you might argue that the model picked up on these details the authors assumed were benign but novel and worked out how to apply them to the problem.
sebzim4500 20 hours ago [-]
I think they are pretty clear that they train on some prompts, given they sell the ability to be excluded
vlovich123 20 hours ago [-]
Unless they know exactly the researcher’s account, they may not know in their end if he had the setting to let them train on his chat logs. They also probably don’t know if he had any correspondence on any forum where he may have discussed this and it got picked up by scrapers.
I’m not saying they didn’t do anything unethical. I’m just saying even if they were ethical, there’s plenty of practical reasons at their scale why a flat out denial is logistically difficult to do
freejazz 19 hours ago [-]
They are already claimed to be liars
baobabKoodaa 8 hours ago [-]
> These responses seem to me to make it abundantly clear who's telling the truth here.
Since it is clear to you, can you articulate what it is? I was unable to infer "the truth" from your message.
hurrrr 7 hours ago [-]
> One can in hindsight see that our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs unforced).
is this true though?
doctorpangloss 16 hours ago [-]
> It would be extraordinarily easy to simply say, this model was not trained on your work, if that were the case.
well, it is trained on their work. all user inputs are paraphrased for training. at openai, at anthropic, at google, and now with all the bedrock models, and at openrouter providers, even if they say zero data retention.
ghshephard 13 hours ago [-]
If this is a "wake up call" - then your legal team needs immediate education.
I don't know how much more clearly they can write:
> When you use our services for individuals such as ChatGPT, Sora, or Operator, we may use your content to train our models.
One of the key selling tactics that companies like Data Bricks or Palantir provides their customers is "Data Governance" - that is, some control over where the data is being used. It's also a reason why enterprises don't use the OpenAI or Anthropic APIs directly - but through secondary sources that have Enterprise Agreements that do their best to make sure that no Company IP is ever retained by a third party, or even exists on a multi-tenant GPU. AWS Bedrock, and companies like together.ai, fireworks.ai have tons of deals that focus very much on data confidentiality.
The reality is - if you want any type of control - you run your own inference, on your own hardware. Anything else and you are at the mercy of third-parties, despite what their contracts might promise you.
pms 13 hours ago [-]
ChatGPT has this option "Improve the model for everyone" in user preferences, which comes with the attached description, meaning that training on user data can be deactivated:
> Allow your content to be used to train our models, which makes ChatGPT better for you and everyone who uses it. We take steps to protect your privacy. Learn more
The "Learn more" link takes you to the link you've shared.
pesacharia 10 hours ago [-]
As I understand it, there is substantial question as to whether that actually stops them training on your data, it just perhaps changes what derivative processes are applied and used.
tpkm 12 hours ago [-]
The training on user data only applies to free accounts - paid and Enterprise accounts guarantee data is not used for training. Plenty of Enterprises use the APIs directly - that's just plain misinformation
carljungslabtek 9 hours ago [-]
This is not true. They claim not to train by default for business and enterprise agreements, but for plus and pro plans they enable it by default and you can allegedly turn it off (I don’t trust them very much though, I’m sure there is something in the T&C saying they can modify that deal any time)
derangedHorse 1 days ago [-]
It sounds like OpenAI is trying to appease the author when they don’t have to by allowing him to rewrite their proof. They probably don’t believe he deserves to, so him asking for a coauthor from Anthropic might overextend their grace in their eyes.
reverius42 21 hours ago [-]
He's not "asking for a coauthor from Anthropic"; he already has a coauthor, who he's already been collaborating with, who happens to also be employed by Anthropic (but whose research in this area is not done as part of their employment at Anthropic).
nezi 24 hours ago [-]
Given that Tristan has said that the proofs that LLMs come up with are mostly "slop" and not up to the standard that human written papers achieve, maybe OpenAI needs an expert like him more than you think to get the result published?
fn-mote 21 hours ago [-]
Almost certainly the Lean proof needs to be decoded for humans and probably also made “human intelligible”.
Now maybe LLMs can also simplify arguments and make sense of them for humans, but we haven’t seen that yet (unaided).
(I haven’t looked at it, personally.)
nl 10 hours ago [-]
It's a Navier Stokes solution.
They don't need any help publishing.
lalalanananana 19 hours ago [-]
Here's a wake up call for everyone sending all of their ip to openai and anthropic. Especially in verticals they intend to dominate. Lol at all the biotech companies all in on Claude and paying millions in fdes creating huge lapses in security as they go.
johnnienaked 15 hours ago [-]
It's too late. Sub models are deployed at every major organization in the United States and all it will take is turning off the option to improve the model for them to train directly on your own personal workflow, which CEOs will greedily eat up instantly if they can reduce labor costs. If they can brute force N-S, automating your finance or SWE job will be trivial. GG to most jobs connected to a computer in the next 5 years.
BTW, this was always the plan from day 1. You will pour all your training and experience into training the model and receive a pink slip as compensation.
JbMaj9 16 hours ago [-]
A mathematician working for the competitor, solving a math problem using our model? Isn’t that the best marketing possible?
Tanjreeve 15 hours ago [-]
If the work done is just "we made other people's work searchable without their consent" it's not quite the same as what they're implying in the marketing of "our model solved this problem".
oliculipolicula 16 hours ago [-]
>only want Buckmaster to be the lead author
Sam+Seb are struggling with their ideological allegiance. This amounts to a confession that there are no reseaechers, only research managers, left at OpenAI. Maybe they even know that they are losing credibility from their main investor(s). They desperately need a domain expert to salvage credibility.
They have no credibility with academia left, obviously, but their main competitor still does. No Millennium prize incoming, I'd wager. For openAI. Let's see mAth get political for once!!
One might be more certain that levent is now going to corner all the institutional support. Go go go!
16 hours ago [-]
segmondy 17 hours ago [-]
For those of us who are into local models and preach it, we are called paranoid. I have often said this, if you are doing any real novel work, or putting your profitable business data/workflow into these models, you're a fool.
davesque 21 hours ago [-]
Honestly this whole thing is so fucking weird. I feel like there's an argument that absolutely no one involved in the final crossing of the finish line to the proof actually did any work (other than just intelligently directing an LLM) and deserves any credit. As the author of this doc mentions, the mathematicians who did the actual work that led to the formulation of this approach (without the use of LLMs; just good ole' fashioned human intellect) are the ones who deserve the credit.
Imagine that a no name janitor used their time in the evenings to go spelunking through the literature to push an LLM to this result. No one would care because that person isn't an anointed expert. So why would the expert deserve any more credit? Because they sort of understand the result, even if they couldn't have achieved it on their own? The whole issue of credit for AI-assisted discoveries seems like it's going to run into a brick wall pretty soon.
GPerson 20 hours ago [-]
I agree for most of the people in the story except Buckmaster himself seems to have been supplying real ideas.
dudeinjapan 19 hours ago [-]
Sounds like the plot for Good Will Hunting 2.
dd8601fn 18 hours ago [-]
Good Will Hunting 2: Hunting Season is taken.
It’ll have to be Good Will Hunting 3.
bitwize 15 hours ago [-]
The Hunting for More Money
johnnienaked 15 hours ago [-]
Token Bill Hunting
WD-42 18 hours ago [-]
Yup! I wanted to side with the mathematician on this one but I read the statement only to discover that they were also pushing an llm on someone else’s idea producing mountains of slop.
Have LLMs actually improved anything? Is mathematics better off than if these slop proofs didn’t exist? Who or what is actually benefiting here.
whateverboat 8 hours ago [-]
> Now that we can see their work, the approaches appear to be different. It is also worth noting that our latest model can solve many, many other math problems.
Is that a threat?
16 hours ago [-]
aaron695 20 hours ago [-]
[dead]
colinhb 1 days ago [-]
I think it's very unlikely that Tristan is making up these quotes, or pulling them out of context:
> I said that if OpenAI released its result in the way proposed I would go
public with what happened. The reply was, “Why would you ruin your career?”
I replied that I am an academic, and asked why he thought going public would
ruin my career. The reply was, “If you don’t want me to be nice, then I don’t
have to be nice.”
Whether and how OpenAI's work on this problem was contaminated by knowledge of Tristan and Levent's work is tangential to OpenAI bullying other researchers into adopting their narrative and dissociating with dis-favored collaborators (ie Levent at Anthropic). Though the latter behavior (threats, intimidation) may weigh against OpenAI in trying to understand the former issue (contamination).
CamperBob2 1 days ago [-]
>I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
If this is true he should release the actual emails. This is a very serious accusation and he shouldn't demand that the reader judge it on hearsay.
magicalist 1 days ago [-]
> If this is true he should release the actual emails
these were statements while on a call, and at least the career comment Bubeck has admitted to while doing damage control ("I deeply apologize for this extremely poor choice of words, it is the opposite of what I was trying to convey. (I should say that I retracted them on the spot by the way.)"[1]).
What a horrible response from Bubeck. How did he think that would make him and OpenAI look good to tweet that?
adastra22 22 hours ago [-]
Did we read the same tweet? It felt very forthright and level-headed to me. Not at all what I expected.
geocar 4 hours ago [-]
He seemed to know a lot about what was going on in the state of the art despite not being a fluid dynamics expert or having any in “the project”, and absolutely nothing about what his own employees did with Tristan.
The most uncomfortable piece is where he shows screenshots “proving” his earnestness is unrewarded instead of the chats that are actually being complained about.
That he does so with tremendous gymnastics is perhaps soothing to investors, but every researcher I show this post to today says “see I knew ‘AI’ would steal my research”
So yeah, maybe we are reading different things.
fn-mote 21 hours ago [-]
> Anthropic models had been used in their proof of Euler blowup; I therefore felt I could not consider Levent to be an independent academic
So using someone’s models makes someone who works for the competitor not “independent”? When their coauthor is? What does that even mean?
I almost stopped reading this extra long post entirely at that point.
This is not a good look in my book.
djokkataja 20 hours ago [-]
The first part of the sentence is important:
> Importantly it was admitted that internal Anthropic models had been used in their proof of Euler blowup; I therefore felt I could not consider Levent to be an independent academic.
If an Anthropic employee is doing independent research, but with models that aren't available to the public (because they're internal models), then . . . idk. It's not clear to me why that should necessarily require a refusal to cooperate between OpenAI and Anthropic employees who are excited about solving a problem like this.
For me, the bigger question here is what "internal models" means to these employees, especially in the context of the OpenAI employees repeatedly avoiding directly answering whether their model had been trained on Tristan's and Levent's ongoing work on the problem. It had always seemed like a loophole that AI companies might be tempted to exploit: yeah, they can say that they won't train on your data, but if an AI company doesn't care about ethics, they might go ahead and train a model for internal use only on everyone's data anyway, just to have as much data as possible and potentially gain an advantage in what the company can internally do. They could never publicly release any versions of a model like that, of course. And of course this is speculation.
adastra22 20 hours ago [-]
This is being reported as OpenAI wanting to strip an Anthropic employee of academic credit for the work they did. What the OpenAI person involved is claiming is that they wanted the outside researcher(s) to put their name on OpenAI's work: to headline OpenAI's publication of what they earnestly believed to be an independent result.
If true, that's generous and beyond the level of generosity one should expect. Extending that courtesy (beyond academic norms) to a competitor is expecting too much. It take a result OpenAI spent millions of dollars on, and put "Anthropic Researcher" right on the cover.
This is, of course, taking OpenAI's side of the story at face value. But it is a consistent, coherent, and ethically justifiable series of events, if indeed it happened that way.
tovej 15 hours ago [-]
Nothing ethically justifiable about it, even if you take their word for it.
If OpenAI thinks someone should be the lead author for a paper, that person should have full discretion to decide who the co-authors are.
Even if the co-author's contributions were non-technical
So no, there's no world where you can ethically extend the right to publish a result and decide who the authors are from the outside.
johnnienaked 15 hours ago [-]
It's a ridiculous explanation that makes no sense.
troupo 13 hours ago [-]
> What the OpenAI person involved is claiming is that they wanted the outside researcher(s) to put their name on OpenAI's work
> If true, that's generous and beyond the level of generosity one should expect
"We highly likely stole your work, and threatened you with 'this is bad for your career' and we refuse to acknowledge any work by your collaborator just because he works at a competitor, but we are so so so so generous"
adastra22 6 hours ago [-]
The two proofs are structurally very different, and don’t even prove the same conjecture. It’s becoming very clear that OpenAI did not steal anything here.
geocar 4 hours ago [-]
Bullshit.
Ain’t no way you read one 100 page paper let alone two with enough understanding to make such a claim.
You ought to be ashamed of yourself.
CamperBob2 3 hours ago [-]
User name checks out
troupo 5 hours ago [-]
If it was "very clear", OpenAI wouldn't be threatening the researcher with "it's very bad for your career" or state "we can't tell you if it was trained on user input".
Additionally, according to the researcher, OpenAI's proof follows the same approach they used, and which was largely unused in academia, but OpenAI claimes they arrived at it immediately.
huhtenberg 9 hours ago [-]
Skipped an important word, didn't you?
> internal Anthropic models had been used in their proof ...
retsibsi 13 hours ago [-]
I appreciate that he responded with (seeming) openness and detail, rather than just posting some pithy insult or whatever would have won him the twitter battle, but this part feels like serious gaslighting or, at best, self-delusion:
> Genuinely, at that moment, I was trying to care for him and do a last ditch attempt to get a chance to give them all the credits that they deserve.
The allegation he is responding to, and which he does not seem to have disputed, is the following passage from Buckmaster's statement (https://cims.nyu.edu/~tristanb/statement.pdf):
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
freejazz 19 hours ago [-]
>I refuted all these accusations but he replied “there is nothing you can do, I simply do not trust you”. I was confused why one would turn an incredible source for celebration (of their achievements!) into such bickering
Pretty insane if he couldn't figure out why there would be bickering in this scenario...
johnnienaked 15 hours ago [-]
I especially loved the part where he claims they spent $60m on compute to push on N-S because of a Twitter rumor, and they totally didn't steal the idea from mathematicians using their tools.
I really don’t understand which party you are referring to.
CamperBob2 5 minutes ago [-]
Anyone who comes close to solving a Millennium problem can honestly say they are on the cusp of greatness. Not sure what this question proves either way.
sinuhe69 1 days ago [-]
No wonder - they're all working for ant! Birds of a feather flock together.
johnnienaked 15 hours ago [-]
Sounds like a great time for open mathematical research is upon us. /s
NewEntryHN 5 hours ago [-]
This is hard to argue without a fine understanding of how much insight OpenAI had about the stab at the problem from the "public rumor" alone.
If there was any sort of coarse insight that "they're trying to solve it this way", then both those things can be true:
- The massive amount of compute from OpenAI re-discovering Buckmaster's work solely from the coarse insight (and solving the rest as well).
- OpenAI still acknowledging they basically scooped the coarse insight using compute, and they're willing to credit Buckmaster.
Buckmaster says "Concretely, what Levent and I did was to take the Cordoba and Martinez-Zoroa program, which achieved blowup results with rough forcing, and, with a great deal of help from LLMs, push it to smooth forcing and to the incompressible Euler equations".
Could that simply be the prompt they used at OpenAI? How "stolen" would the proof be in that case?
mentalgear 18 hours ago [-]
A wake up call for anyone using (openAI) chatbots : your data, ideas and execution can become their spontaneous 'inspirations' at any time - even if the LLM providers are 'just' using a meta concept monitoring system across all incoming user-data, running in the background constantly checking for 'lift-ables' for their company's bottom line.
JumpCrisscross 17 hours ago [-]
> These two bullet points are extremely suspicious if you were honest
I don’t think we can consider these accusations separately from the evidence being unveiled about OpenAI’s culture by Apple’s lawsuit. These guys seem to openly embrace the strongest interpretations of “good artists copy, great artists steal.”
dgellow 1 days ago [-]
If that part of the PDF is true that’s so disgusting, psychopathic behaviour
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.” Some time later Levent received a text proposing that he and Sebastien speak one on one, saying, “I don’t know if Tristan is being fully rational right now.”
27183 6 hours ago [-]
This kind of coercive, threatening rhetoric is really not surprising at all. This is how tech companies operate.
What stands out to me is the naive lack of operational security on the part of academics, who should know better than to touch this SaaS crap with a ten foot pole.
swozey 4 hours ago [-]
I'm imagining all of the potential targeted customer lists now. Grab all those sweet .edu, et al logs without anyone thinking the wiser- until you release a stolen solution. What've they got on .gov?
AnimalMuppet 1 days ago [-]
With "rational" being defined as "not making us look bad" at best, and "letting us take credit for their work" at worst.
Yes, that is disgusting.
5 hours ago [-]
16 hours ago [-]
viccis 1 days ago [-]
>After hearing of the rumor, OpenAI started researching Navier Stokes with a new internal model.
This is the most suspicious thing to me. If their chat data were available to the corpus to be trained on (I thought they claimed not to do this?) then it really might be as simple as querying the model with "describe recent work from Tristan Buckmaster" and it will spit out this problem and his approach. No need to directly read his user data.
This is basically just scooping, real scumbag behavior.
rakejake 1 days ago [-]
OAI doesn't need to mention Buckmaster's name directly in a prompt. They just need to select a basket of sessions that is guaranteed to contain Buckmaster's and then direct the LLM to attack only a specific method/angle. This is trivial to do while maintaining plausible deniability about not using his work.
spongebobstoes 21 hours ago [-]
what reason do we have to believe that they did this? both things were proved by AI, isn't it logical that they could have very similar approaches?
it is common that multiple people essentially simultaneously prove/invent the same thing
I see zero evidence of wrongdoing
rakejake 18 hours ago [-]
OAI started working on this only after they found out it was close to being solved. They threw a team of researchers who spent sleepless nights + a ton of compute. This is not exactly healthy academic competition - it's like if you spend a year hunting for oil fields and finally find a very promising area to be explored, only to find that Exxon tapped their entire exploration unit to go all in and and find it overnight just to stake claim to the discovery. Tao said it right - math should not be treated as a non-renewable resource to be mined.
spongebobstoes 17 hours ago [-]
that is not related to the accusation that they literally stole Buckmaster's work, which seems baseless
I don't agree with the oil claim analogy. this is knowledge, freely given to the world. not something hoarded by a corporation
JumpCrisscross 16 hours ago [-]
> this is knowledge, freely given to the world
Even if you’re starting from a position that credit for a discovery literally can’t be stolen, that still doesn’t resolve in OpenAI’s favor here.
spongebobstoes 16 hours ago [-]
it seems like OAI tried to share, but didn't want to share with an Ant employee. a bit childish, but understandable to want to avoid a headline "Anthropic researcher solves Millennium problem"
it seems like Buckmaster got one-upped and is upset. understandable, but I find their reaction childish as well
JumpCrisscross 15 hours ago [-]
> OAI tried to share, but didn't want to share with an Ant employee
Why does OpenAI get to dictate who Buckmaster can claim co-authorship with?
> I find their reaction childish
OpenAI may have, with full plausible deniability, taken Buckmaster’s work and passed it off—in substantial part—as their own. (Fitting into a fact pattern of them having tried to do the same with Apple.)
There is a material takeaway for anyone who does creative or otherwise unique work from this. (Which is unfortunate. Whatever happened here, AI clearly accelerated the discovery process.) For anyone else, I agree it’s just drama.
spongebobstoes 7 hours ago [-]
OAI didn't try to claim Buckmaster's work as their own. OAI tried to let Buckmaster present OAI's work. it is nearly the opposite of your accusation
s1artibartfast 53 minutes ago [-]
Pretty strange if they didnt steal Buckmaster's work, right
oliculipolicula 6 hours ago [-]
More like OAI needed someone like Buckmaster's stamp of approval. They know they can't get Tao's. Now they won't be able to get Zoroa's. You have to hope the rest of the guys they didn't know to cite will readily give up honour and dignity
It doesn't seem like there's anyone at OAI who knows much about the problem they are solving. Seems like CS theorists or algebraists* trying to own the pros by driving a car that's beyond their skill level. For one, they didn't cite the guys that B&A based their work on.
*It would be most fair to say there are no analysts on board, nor are they likely hire any soon; those are the least impressionable people in math. Applied math PDE elves who hadn't already left on the world-model boats would have jumped off around the time that eg Ilya did because they wouldnt have been able to stand the three Bs pretending to be experts in fields they imagine to be "adjacent".. like Public Relations
tripzilch 11 hours ago [-]
but a very large part of the whole model was trained on work in a manner the authors didn't consent to, the "for research purposes only" datasets of the entire Internet, etc
and you can argue this is "fair use" or whatever, not the point now, the point is that it definitely makes those accusations no longer "baseless".
in addition, it is not given freely to the world, it is the knowledge of the Internet/WWW being sold back to you as a subscription service. it's not free. and it's not even "given", because they can (technically) turn off the tap at any moment and you don't have it any more.
KoolKat23 6 hours ago [-]
Humans do this all the time. You watch a YouTube video and subconsciously choose the same colour palette. They hear a rumour that it's a solved problem, i.e. they were pointed in the right direction that is all.
I mean some companies glean insight into new products merely by asking other people what they do for a living.
People need to get over this ownership thing, it's being taken too far. Humans benefit from the efforts of others simple as that.
JumpCrisscross 17 hours ago [-]
> what reason do we have to believe that they did this?
The culture at OpenAI being systematically revealed by Apple’s lawsuit, for one.
j16sdiz 20 hours ago [-]
This depends on what "proved by AI" meant.
Was that a one shot prompt? or something guided by human, step by step?
If that's the later, it won't use the same approach when not guided by the same human.
efavdb 7 hours ago [-]
Apparently Tristan was working on this problem for years, with AI providing help. Vs OpenAI spending a week with their new model, potentially having access to Tristan’s work.
lern_too_spel 16 hours ago [-]
Then they say that they aren't sure if their model accessed the other researchers' private data. Why not wait until they know for sure, rerun in a way they can ensure doesn't access the other researchers' private data, or wait until the other researchers have published to make sure they're not stealing another person's work to build upon?
epistasis 1 days ago [-]
> OpenAI says the model didn't access his user data directly, but leaves unanswered whether Tristan's chat conversations were part of the training.
This sort of cagey half-answer is highly suspicious and indicates that yes OpenAI did actually "access user data directly" because they are only willing to say that the "model did not access user data." That has a very specific meaning, the model looking up user chats, that they can defend.
So, everything we submit to OpenAI can be considered to be part of future models, right?
rzmmm 15 hours ago [-]
Isn't it obvious? Web scraping and even scanning written books is at record levels because the data is so useful for training. They are using every byte of user data.
chinathrow 12 hours ago [-]
If I were involved, I'd file suit and have them to at least not delete any proof.
kccqzy 18 hours ago [-]
OpenAI says
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models
That basically means, we don’t know, and we hope the model didn’t look up user conversations, and the best thing we can do is hope.
That’s seriously disgusting. I can understand why on a technical level why perhaps it is impossible to answer what exactly the model had access to, but it still is disgusting.
doctorpangloss 16 hours ago [-]
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models
it just means that they've been working with paraphrased user data everywhere, which any smart person can figure out is how anthropic and openai train on so called non-retained data.
jimbob45 13 hours ago [-]
How could they possibly know? If Tristan posted on r/math and they slurped that up as training data, that would count, no? They might never even know. I can’t envision any absolute statement by them claiming that they didn’t use his work that survives legal rigor. That is, this statement was never not going to be in this post in any of the infinite multiverses.
kccqzy 7 hours ago [-]
This is private work that is being done with Codex. Not drafts shared on r/math. I suggest you reread the article.
This is more of an expectation of privacy issue. If you write on an envelope and USPS has a copy, you have no reason to be mad; if you write on a letter inside the envelope and USPS still has a copy, you could rightfully be mad and say this is disgusting.
mahart 17 hours ago [-]
Every other paper in existence has been ingested with 99% of writers not knowing it will be retroactively used for training. But session data which is disclosed as being used in terms of service is disgusting?
If the work is duplicative/derivative then the preprints they put in sessions can be shown by the users and we can see.
epistasis 5 hours ago [-]
What's disgusting is that they refuse to acknowledge what they and only they know: were these chats fed in to the new model that found these results?
Nobody would be disgusted at following stated policy, that's doesn't make sense, without first objecting to the policy. But you're bringing up distractions from the actual concerns: did OpenAI use the private chats and why won't they confirm or deny it?
bawolff 1 days ago [-]
> OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic
well that sounds like an asshole move.
mti 1 days ago [-]
In a serious discipline like mathematics, this isn't just an asshole move, but a career-ending level of academic misconduct.
augment_me 1 days ago [-]
*In the old discipline of mathematics.
We are in new times, where capital and compute decides mathematics, so we don't give a shit anymore
cindyllm 1 days ago [-]
[dead]
bawolff 1 days ago [-]
I wonder what is required for it to cross the line into criminal blackmail.
dgellow 23 hours ago [-]
From the company that likely committed federal crimes by hacking Hugginface
numpad0 21 hours ago [-]
One thing that wasn't obvious to me or adults around me when I was younger: most laws define whatever acts a law punishes as individuals commiting to it, not as situations manifesting anyhow. It's not a murder just because someome died hit by a bullet you fired, but you have to have personally decided to kill that person leading to their death[1][2].
OpenAI's LLMs are not humans, and neither is the company. So by this logic, I think there's a chance that nobody committed a crime by hacking Huggingface, and also the chance that a lot of military and police organizational orders become illegal if OAI's doings would be illegal.
IANAL and all I have is a bucket of popcorns, though.
1: not a meaningful defense in a real trial, also gross negligence exists
2: this also explains insanity defense; if you were so out of your mind that you could not have held such a thought, it is considered out of scope for justice systems
bawolff 20 hours ago [-]
> OpenAI's LLMs are not humans
Neither are guns. Which is why we punish the person shooting the gun and not the gun.
Industrial equipment, which is how i would classify LLMs, hurting people is nothing new. The relevant questions are:
- did someone intend it to happen?
- was someone negligent in taking reasonable steps to prevent something foreseeable?
The justice system doesn't punish people for legitimate accidents. e.g. if you are shooting at a shooting range, take all reasonable precautions, but someone was hiding behind the target, you are probably not guilty even if you shoot the guy.
As far as openAI goes, the logic is the same. The question is, was it intentional, was it unintentional but reasonable precautions weren't taken or was it truly an accident?
numpad0 2 hours ago [-]
I mean, they are going to be at least culpable for gross negligence, but who made the decision might become an interesting point of discussions(mainly in deciding precisely how they should be punched in their face and less in whether they should be at all).
ted_dunning 14 hours ago [-]
And the third time it happens, is that still just an accident?
fn-mote 21 hours ago [-]
> a career-ending level of academic misconduct
There is no such thing anymore.
Falsified data and published? Absolutely no problem. Keep your tenure.
It’s even hard to lose your position as president of a university due to egregious misconduct.
asey 22 hours ago [-]
That's not quite right - Levent and Buckmaster did not actually have the Millennium Prize qualifying NS solution but something more limited. OAI invited Buckmaster to join and help rewrite the full solution paper but did not feel it was appropriate to invite an Anthropic employee to join as well - particularly given they were using internal unreleased models.
golem14 12 hours ago [-]
Something like this got Schmidt/ Jobs into serious trouble …
matt3210 16 hours ago [-]
Agreeing to partially credit is an admission that they used his work.
rsrsrs86 19 hours ago [-]
Could we expect otherwise from big tech?
yk 21 hours ago [-]
The last two points are disputed/sound significantly more reasonable in [0]. So from what I gather, Buckmaster realizes sometime during the call that the biggest result of his career is going to get steamrolled (the blowup of Navier Stokes is a much bigger deal than the blowup of 3D Euler), and on the other hand the openAi guys realize that they are basically talking about internal results with Anthropic and probably have to call corporate right after this call. Between these two stressor the conversation appears to have gone somewhat poorly.
"Since August 28 we have been training a new internal model that has exhibited unprecedented performance in our benchmarks, including mathematics. This model’s training is ongoing and its performance continues to improve."
"When a further trained version of our internal model became available over the course of the effort"
"While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models "
Davidzheng 1 days ago [-]
But Tristan doesn't even want to be credited for the millennium prize--does he? He wanted to be the first to solve it and got scooped (which ig is not a great look for OAI ethically but also not forbidden). And the only reading for her of asking to remove levent must be in light of this offer (to credit him for millennium prize) right? It's not like they can ask him to remove levent from the main papers without this offer(what would Tristan gain?)
My guess is that OAI tried to be "generous" and offered to share credit on millennium with Tristan but not Levent. And Tristan got understandably offended by this offer (which probably oai felt like was the right thing to offer but they couldn't really offer to do the most ethical thing for some reason) and then the random conflicts and weird threats started.
lukewarm707 1 days ago [-]
my reading was that openai did not deny plagiarising the approach from their prompts.
openai then tried to effectively bribe buckmaster with a shared citation, whilst dropping his co-author who works for anthropic.
after buckmaster refused, openai tried to threaten him.
hellohello2 21 hours ago [-]
This definitely feels like the correct reading unless there is information we were not provided with. If the problem were unimportant, there would be no debate that this is not OK...
hellohello2 19 hours ago [-]
I was too slow to edit this post but I'm not sure why I said this. This is too assertive about a situation I don't know much about.
Davidzheng 1 days ago [-]
Oai denies looking at prompts but doesn't deny training on them.
lukewarm707 1 days ago [-]
if they do not deny training on them, they can't deny plagiarism.
fc417fc802 23 hours ago [-]
By that logic everything any LLM spits out is plagiarizing the vast majority of work written prior to a few months ago. That doesn't seem like a useful or desirable line of argument to me.
myrmidon 23 hours ago [-]
Just replace the model with a human student.
"Training" on textbooks => fine
"Training" with unpublished notes from another professor, then publishing something on that exact topic with a similar approach without giving any credit => extremely questionable.
fc417fc802 22 hours ago [-]
Presumably the professor voluntarily provided the notes in this analogy. I think the student would also be expected to cite the textbook if building off of it directly. In contrast, humans are generally not expected to cite "general inspiration" or what have you. So if we're to apply human standards, and assuming that the model was trained on the relevant work, it would only be plagiarism if the model directly built upon that previous work (at least IMO).
The trouble here is that if LLM training constitutes direct use then approximately _everything_ they output is blatant plagiarism, not just a few pieces of academic work.
Conversely if training is viewed as analogous to a student attending classes to learn general concepts (not a perfect analogy, I realize) then nothing they output on their own (as opposed to receiving as part of context) is plagiarism.
Thus this seems like a fairly useless line of argument to me as far as the current topic goes. It either implicates this academic work along with literally everything else or else it does not implicate this academic work. Kind of like nuking an entire city and then saying "mission accomplished, killed the bad guy".
fn-mote 21 hours ago [-]
To be clear, the accusation is that they trained on the chats they used while working on the problem. Not published work or even a preprint.
Your post does not distinguish, and it matters.
fc417fc802 20 hours ago [-]
How does it matter? It either is or is not plagiarism. Ripping off a published textbook isn't somehow better than ripping off private correspondence. Both are serious acts of academic misconduct on account of the part where you knowingly and intentionally portrayed someone else's work as your own.
Note that I am not taking a stance on what openai allegedly did or did not do one way or the other. I am merely pointing out what I see as a fatal flaw in the line of argument presented by the earlier commenter - the idea that training on an item is on its own sufficient to establish plagiarism of it.
lukewarm707 5 hours ago [-]
voluntarily is stretching it for an opt-out
anonymousDan 21 hours ago [-]
This is just a nonsense line of reasoning. Training based on the solution to the problem (or the key insight behind the problem) is clearly a form of plagiarism.
fc417fc802 20 hours ago [-]
What about my line of reasoning is nonsense? I made no claim either in support of or contrary to yours. Rather I pointed out that by this logic literally everything that an LLM spits out is plagiarism of the vast majority of the entire body of human literature in existence. Can you offer meaningful refutation of that observation of mine?
za_creature 18 hours ago [-]
Many do indeed hold the position that all LLM output is uncopyrightable plagiarism. They're probably right, but there's an even stronger argument here:
Science papers of a phd level must contain:
1. one or more novel insights
2. a long list of citations to contextualize them and
3. some work to prove that the insights are in fact meaningful
---
In this context, consider a prompt based diffusion model which, when asked, will happily produce a few pictures of a horse in orbit. You then tell it "silly robot, horses can't breathe in space" to which it adds the necessary space suit in a follow up image.
That image is twice plagiarized:
1. the model did not come up with the original idea of putting a horse in space, nor with insight that horses need a space suit
2. the model failed to cite where it pulled the "horse" and "space" concepts from.
It merely did the work (3) to combine the concepts using the user provided insight.
---
The implied accusation here is that OpenAI used the insights from an existing prompt to train a new model that was able to one shot "a horse race in space" picture, and they were all wearing space suits.
This is still academic plagiarism, even if you disagree that all LLM outputs are.
fc417fc802 12 hours ago [-]
I neither agree nor disagree that all LLM outputs are plagiarism. I merely objected that the line of argument engaged in was specious given the context.
As to your stronger argument. You only cite prior novel insights that you're actively building off of and that (approximately speaking) fall outside of the status quo. You don't for example cite leibniz or newton despite your paper making heavy use of calculus.
So is there any actual evidence that openai trained on the data in question? And further, did the openai proof directly build on someone else's novel insights as opposed to deriving everything from scratch? (I don't pretend to know but the vast majority of what I've seen so far in the comments here is what I'd characterize as brain-dead screeching. Certainly not the level of discussion I come to HN for.)
Separately, consider the implications of what you're arguing for there. Suppose your horse in a space suit picture were somehow valuable to society. Suppose that due to shortcomings of your tool you lacked the ability to readily and accurately identify the originators of the relevant concepts. Should you refrain from publishing this useful work due to the lack of citations? How are you supposed to handle this situation?
Remember that in this analogy everyone throughout society is on the same page that your tool consistently recycles other people's ideas while being technically incapable of producing reliable citations. The question is a simple trolley-esque problem - do you publish without proper citations for everyone's benefit and if so what are you supposed to say?
lukewarm707 5 hours ago [-]
evidence that openai trained on the data: they would have denied it if they didn't train on it.
did the proof build on the insights:
the influence of an individual text in the training data is deeply weighted by quality, relevance, etc. a high quality proof in advanced mathematics written by a codex user is going to get boosted to the max.
the model is post-trained on prompt material. that is again going to boost it.
the prompt will boost this material specifically. perhaps they even rammed dense maths in particular into the model in post training.
anecdotally i have been able to get near-verbatim copies of original material out of models at inference. the type of work that buckmaster and alpoge fed into openai feels like the exact type of concept that would cause an "aha!" or "but what if?" in chain of thought. in fact i would bet that their work is in the logs.
the likes of astra and fable are thought to be up to 10T parameters in size. i consider it highly plausible that a semantic representation of the euler proof could be pulled out of the model weights in good shape.
za_creature 9 hours ago [-]
I am not engaging further. You asked for meaningful refutation of your observation, I provided one.
Academia has stricter rules than regular society.
cerwisc 7 hours ago [-]
The chats the professor had are not generic knowledge. And yes of course you still need to cite Newton and Leibniz depending on what result you want to mention. What’s allowed to be not cited are not status quo, the term you’re looking for is “folklore” results aka results that have been around so long that 1) nobody knows who came up with them or 2) everyone knows who came up with them.
The second point: if you say you can’t prove that OpenAI actually used it, it doesn’t mean that OpenAI did not use it. It’s hacker news not lawyers news here lol. And OpenAI can’t prove that they didn’t use it either. The whole point is that Levent felt he had reasonable suspicion to believe the AI did use the result, because he felt like without his input on an unpublished paper it was unlikely for AI to reach the same result. I haven’t read the paper so I don’t know where I stand on that.
On the last point, about your “for the greater good” argument. It’s higher maths lol. I don’t know about this field but I doubt it’ll be very useful for society. Maybe it’ll make one part 2x faster which makes some rocket cheaper to launch. Does the average person care? Debatable. I think it’s reasonable to hold published papers in proof based fields to a higher standard. Otherwise the current & future problems of ML engineer fields just expand to other fields. No thanks.
Finally, if you anonpost to the autistic Internet forum that everyone else is “brain dead screeching”, it really just says something about yourself lol.
visarga 5 hours ago [-]
Even if AI used the result, AI pushed it to the finish line while Levent and Tristan did not. But I understand the approach was different, the information leak was only that it was "doable".
cerwisc 3 hours ago [-]
The approach was the same.
Tanjreeve 15 hours ago [-]
That's exactly the argument of the people calling it plagiarism machines. No-one ever really did refute it there was just a bunch of settlements for elite institutions so they weren't left empty handed like the various small time creators/authors etc were.
I think the bigger issue here is this feels like some PR smoothing happening that after all the work that went into "it's safe to use for enterprises" now we have what looks like openAI using private user data to scoop novel research and the question of why couldn't they do it for an enterprise with much more money on the line.
fc417fc802 12 hours ago [-]
> now we have what looks like openAI using private user data to scoop novel research
Is there any actual evidence of that? All I've seen so far are empty accusations because "it would be in their interests" or whatever. Personally I'm inclined to believe that they honor their terms until it's demonstrated otherwise.
lukewarm707 5 hours ago [-]
their terms grant them an irrevocable license to your data unless you specifically opt out of it.
freejazz 18 hours ago [-]
What does it matter? We're supposed to not call it plagiarism anymore because it's inconvenient to call it the plagiarism machine? What's your actual argument? Otherwise it's completely irrelevant what an LLM does in other contexts or what we call it
hellohello2 21 hours ago [-]
This is a common misconception, so its understandable that you have it. Generative models can both plagiarize and generalize. The question here is which of the two happened.
fc417fc802 20 hours ago [-]
A needlessly condescending tone while failing to address the topic at hand. The person I replied to advanced the claim that training was sufficient to constitute plagiarism. You appear to be claiming that it is possible to generalize instead of plagiarize after training on something, so I take it that you must necessarily disagree with the original claim?
hellohello2 19 hours ago [-]
What I meant to say is that, in many cases, a generative model's output is not in fact steered by minor amounts by lots of training samples, but instead steered by a just few samples. Some outputs are influenced by many inputs, and some by very few, it really depends.
In answer to a post suggesting that training on a datapoint could mean plagiarism, you said that this would imply that all outputs are plagiarized. This is not the case, no, because generative models do not "copy" or "create", they do both at different times.
I did not agree or disagree with the original poster, I was explaining to you why I thought you disagreed with them. If you understand what I said above, then why do you disagree with them?
EDIT: I just saw your other post on "general inspiration" and I believe I read the situation exactly; you appear to believe that inputs used to train generative models get "lost in the parameter soup", but it is not always the case.
fc417fc802 11 hours ago [-]
> In answer to a post suggesting that training on a datapoint could mean plagiarism, you said that this would imply that all outputs are plagiarized.
We read the original differently. As clearly stated in my previous reply to you, I interpret it as claiming that all outputs are necessarily plagiarizations of the training data. That is not my claim (as you wrongly stated) rather it is the claim I am responding to. I observe that it is absurd to object to a single action being a transgression on the basis of an argument which implies that all actions are inherently transgressions. Notice that nowhere do I take a position on whether or not the argument about all actions being transgressions is true or false.
> you appear to believe that ...
I do not, no. I have not taken a position of my own here. I've merely objected that the one I responded to does not make for a sensible line of argument in context. It seems that you (and many others) have read my objection to position A as support for position B and attempted to infer what I think from that.
hellohello2 3 hours ago [-]
I simply do not see how you can interpret "if they do not deny training on them, they can't deny plagiarism" as "all outputs are necessarily plagiarizations of the training data".
There is a difference between claiming an action is a transgression, and claiming it could be one.
demibabs 23 hours ago [-]
Isn’t that one of the most salient and straightforward argument against LLMs?
persedes 22 hours ago [-]
Just overfit ad infinitum:)
lukewarm707 22 hours ago [-]
as good academic conduct you may cite the source of the work you are quoting or paraphrasing.
as bad academic conduct you may steal someone else's unpublished work, work on it yourself for a bit, and then publish it as your own work. and then threaten the original author!
sdenton4 1 days ago [-]
At this point who knows? Maybe the agents got into the user data while no one was looking.
mswphd 24 hours ago [-]
they're using a new model trained since the prompts happened. They are not denying the other group's solution may have been in their model weights, despite it being unreleased.
itemize123 18 hours ago [-]
I think it's worse - openai didn't even deny training on their exact manuscript (when they most likely opted out);
WinstonSmith84 14 hours ago [-]
Well, regardless of the drama, the interesting part on Anthropic vs OpenAI is:
- One of the two main persons work at Anthropic and "almost" or "partially" solved the issue, but eventually didn't succeed
- An external person with just an excerpt of the chat and certainly less versed in Mathematics (than these 2) tackled the problem.
There is little doubt that OpenAI is so much ahead and maybe the gap is even larger than what we see on Astra vs Fable.
pfbtgom 14 hours ago [-]
I recommend that you read the linked PDF before drawing any conclusions.
This is sort of a weird interpretation:
> One of the two main persons work at Anthropic and "almost" or "partially" solved the issue, but eventually didn't succeed
- It wasn't an Anthropic endorsed effort.
- Solving this class of problem means a march of progress A -> B -> C -> D. If a student turns in a test that jumps from A -> D without showing any work they're either brilliant or cheating (probably cheating). Further, each step of progress isn't the same proportion of effort. What if moving from C -> D was actually the smallest contribution and just required a novel perspective to make the breakthrough.
This part is wrong:
> An external person with just an excerpt of the chat and certainly less versed in Mathematics (than these 2) tackled the problem.
- it was a whole team at OpenAI working on the problem
- it wasn't a chat excerpt, it was more like their entire git repo and project progress reports
grey-area 14 hours ago [-]
Often with a hint on how to solve something, solving it is much much easier. It seems like that's what happened here.
Very weird behaviour from OpenAI, offering partial credit to on person, but not the other person involved. Trying to bully the mathematicians involved (see threats quoted upthread).
I suppose it's the sort of amoral behaviour we've come to expect from them.
HWR_14 16 hours ago [-]
I don't understand the desire to remove Levent. "Off the clock, when Anthropic engineers want to break new ground, they use ChatGPT" sounds like a great ad.
menato 6 hours ago [-]
All these small details will be forgotten in less than a year. But association "Navier-Stokes -- OpenAI" will be part of the history. In my opinion, they are thinking about long term PR here.
wraptile 18 hours ago [-]
> but only if they remove Levent as an author, as he works for Anthropic.
What a terrible look for OpenAI to die on such a tiny hill right there. I wonder whether Anthropic would have made the same requirement.
mepiethree 16 hours ago [-]
well they didn't really die on the hill. the NYT headline is still "OpenAI says it has cracked one of math's millennium problems" and the whole article doesn't mention this controversy. Normies don't know what NS is and don't care, they'll just see "wow OpenAI is the best I guess"
modeless 1 days ago [-]
> After hearing of the rumor, OpenAI started researching Navier Stokes with a new internal model.
I don't think this part is accurate. OpenAI was researching Navier Stokes before. It's possible that they started on a new approach after hearing of Tristan's success, however that is not proven and I expect we will hear OpenAI's side of the story today.
loose-cannon 1 days ago [-]
In the PDF:
"The route to the Clay problem through a smooth force, options c and d in Fefferman’s statement of the problem, is the route Luis and Diego opened and the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it. It is not the direction one arrives at in a few days by giving a model the problem statement. When I heard “forced,” it was a bright red flag."
Whether or not they were researching it before isn't the concern.
treis 1 days ago [-]
This seems unsupported. OpenAI has access to internal models that the general public doesn't have and a compute budget that dwarfs what an NYU professor would have.
mswphd 24 hours ago [-]
it's very possible they only had to use the massive compute budget because they were trying to plagiarize his work before he published it though, e.g. autonomously do things in ~7 days what he had likely been thinking about for ~1 year.
treis 23 hours ago [-]
This doesn't really make sense. You don't need massive amounts of computing to plagiarize something.
The most nefarious explanation seems to be that they got wind it was possible to solve NS via LLMs and perhaps a small nudge in the right direction.
fn-mote 21 hours ago [-]
> You don't need massive amounts of computing to plagiarize something
The compute was used to leapfrog the human team, using their ideas and pushing them to a solution of the general problem.
Plagiarism isn’t being used in the literal sense.
vlovich123 22 hours ago [-]
Of course you do if you're a) only given a partial solution b) racing against someone else using a competing AI.
The open question was whether their LLM got the nudge in the right direction because it got access to the chat somehow (e.g. automated training that scraped his chat logs) or just a high level "Navier stokes can be solved through LLM". It sounds like the former may have happened although right now we just have an accusation and a weak denial.
devindotcom 20 hours ago [-]
wasn't it reported elsewhere that they used the equivalent of $22M (street) in Astra tokens? obviously it's not the same when you own the machinery but still.
n2d4 1 days ago [-]
You're right — the wording in the doc is that the "first prompt" was sent after learning about the rumor, although this might be the first prompt of this solution approach, not necessarily first prompt to any Navier Stokes solution.
1 days ago [-]
lern_too_spel 15 hours ago [-]
Sama himself has now confirmed OpenAI started researching NS after the rumor.
"It is true that we tried this because there were rumors on the internet last week that Anthropic's models had solved a millennium problem and we were curious if ours could do it too."
It seems exceptionally unlikely to me that OpenAI would be "reading user prompts". More likely is a leak somewhere else. Obviously Anthropic/OpenAI are engaged in espionage stuff with each other. I am guessing Levent just mentioned something to someone at Anthropic and it got out.
rockdoe 1 days ago [-]
Of course they are "reading user prompts" in the sense that it's being used to further training. That's why you have to pay up to opt out of that.
That's also why they refused to answer that question: the answer is obviously "obviously"!
lukewarm707 1 days ago [-]
if training is on, they are reading user prompts. training is on by default.
simianwords 1 days ago [-]
why is this downvoted? do people really think OpenAI is snooping at people specifically? like they are looking for good leads into new problems or ideas and they found this guy's codex thread and used it? come on man, even for conspiracy theories this is stupid.
They explicitly say that they use your "content" to improve their models. Considering they practically have infinite compute at their disposal, why is it surprising that they would look for juicy data in there to make them look good ? When they ingested basically the entirety of human knowledge without regard to the rights of others, when they burn books by the thousands, when their relentless barrage of bots have rendered the Web borderline unusable, why would they stop at that line ?
simianwords 1 days ago [-]
[flagged]
GPerson 1 days ago [-]
You’re always going around insulting people. Does it make you feel powerful?
Is it really hard to believe the people who would consume all the world’s data regardless of copyright and norms and permissions would not respect the data privacy of a user?
gcr 1 days ago [-]
this is a highly marketable problem and specific teams at OpenAI were aware of specific competitor efforts. I would be surprised if this were happening on a large scale, but
1. Less than a hundred people in the world are working at this problem,
2. A significant fraction of those happen to work at competing hyperscalers,
3. Those hyperscalers repeatedly show themselves not to take user privacy seriously
simianwords 1 days ago [-]
I'm not sure what your points 1 and 2 have to do with anything. Both directionally increase the probability of hyperscalers also finding the solution independently.
> Those hyperscalers repeatedly show themselves not to take user privacy seriously
where? Any examples?
hn_acc1 1 days ago [-]
Have you been living under a rock? None of the big tech companies care one bit about user privacy in the US.
CamperBob2 1 days ago [-]
That's not an answer.
hellohello2 21 hours ago [-]
Any degree of tracking what people do is unprivate. Every single web interaction you perform is tracked. All LLM companies store all your conversations by default. Do you need more examples?
freejazz 18 hours ago [-]
It's one of those questions that sets the bar so low, the conversation might as well not be taking place
hn_acc1 1 days ago [-]
Even if they claim not to use it, they're probably using it and hoping they don't get caught. They have zero ethics or morals, they just want to "win" to get mega-rich.
jawilson2 1 days ago [-]
> do people really think OpenAI is snooping at people specifically
yes.
Rules and laws are for the poor.
dgellow 23 hours ago [-]
Obviously, yes
freejazz 18 hours ago [-]
> do people really think OpenAI is snooping at people specifically?
I'd be shocked if they weren't
HDThoreaun 1 days ago [-]
Yes that sounds like something openAI would do to me. Not that they’re just looking through random professors chats but they heard buckmaster made progress on navier stokes and decided to read his chats.
Panoramix 1 days ago [-]
Yes, that would be par course with OpenAI behavior
Betelbuddy 1 days ago [-]
In other words...they should have used Bedrock...
woadwarrior01 12 hours ago [-]
> - Levent works at Anthropic, but this research was independent of his work there, with a mix of GPT and Claude models.
One minor wrinkle in Alpöge's narrative, he refuses to deny that he hasn't used any non-public Anthropic models for his "independent" research.
people are so greedy to have their name on something forgetting a) human progress is shared through its cycle. people build upon eachothers ideas and knowledge. and b) all these models consist of information stolen from others. Who really solved it if all u did was prompt some illegally obtained repository of data..??
Its like saying you won the car race, but you stole the fastest car and had your m8 drive you around the track.
ncr100 9 hours ago [-]
To build on this, is it correct to say ANY training on data from any individual originating human deserves to be cited and attributed on the solut6paper?
I wonder whether these bickering users (of AI various products) with math degrees are missing the forest for the trees.
nautilus12 7 hours ago [-]
This should be a watershed moment for everyone using a commercial LLM. EVERY PIECE OF INTELLECTUAL PROPERTY YOU DISCUSS WITH IT BECOMES THEIR PROPERTY. FULL STOP.
It is impossible to fight against this. It is statistically obfuscated, they can just argue it is a derived work.
m00x 21 hours ago [-]
I haven't seen any proof that OpenAI asked Tristan to remove Sebastian from the prize. Until we have proof of this, it would be wise to offer conclusions.
Same for NS validity. This was not validated by the community yet.
mepiethree 16 hours ago [-]
what would such proof look like? It's not like allegations are written in Lean
applicative 20 hours ago [-]
you can read it in buckmaster's document.
wbl 7 hours ago [-]
And OpenAI confirmed it!
someothherguyy 20 hours ago [-]
its what all LLM users have all been doing, profiting off others' IP through a number cruncher, while relinquishing their own
1 days ago [-]
lukewarm707 1 days ago [-]
please could you change 'drama' and 'accusation' to something more formal like 'allegation'.
the paper makes a very serious allegation of dishonesty and possible academic misconduct.
the governance and integrity of openai is of importance to the welfare of society. this is not a matter of drama.
ryan_n 1 days ago [-]
The [lack of] integrity of OpenAI (and any other frontier lab) should already be pretty solidified. Among other horrible things, these companies stole millions of IPs and no one seems to care anymore. Regardless of what you think of the product they are making and the success of ai/its impact on humanity, these companies objectively do not have much integrity.
simonw 1 days ago [-]
How do you feel about the integrity of the machine learning researchers over the past twenty years who trained models on scraped internet data that weren't particularly powerful and didn't attract any attention?
ryan_n 1 days ago [-]
If they scraped internet data in the same way as current day frontier labs do, then I feel the same exact way about them. Why would I feel any different if that is the case?
simonw 1 days ago [-]
My point is that researchers and academics really have been doing this for decades - it's the reason projects like Common Crawl and LAION exist.
I think it's notable that nobody was calling out those researchers for their lack of integrity, because the systems they were building did not seem like a threat to anyone.
OpenAI etc get accused of a lack of integrity on this precisely because the systems they are building work, and are profitable.
My personal opinion here is that integrity is more about what you build with the data. I think saying "scraping means you lack integrity" is a simplification.
ryan_n 1 days ago [-]
You're right, it was an over simplification. I think public exchange of data is great for innovation and research (Common Crawl/LAION). But I still think scraping proprietary data without consent or attribution is generally bad (also Common Crawl/LAION).
Then you have OpenAI etc.. who build these multi-billion (trillion??) dollar machines and sell them back to people, using everyone's proprietary data, and (among other things) tell everyone it's going to take their jobs. That combination of things doesn't scream integrity to me.
Still, it's undeniable that these machines could be beneficial for humanity (cancer research and such). So, I'm sure many people would say the good out-ways the bad. I don't know. Seems that would set a risky precedent for future companies, but maybe not.
smcg 1 days ago [-]
you massively collapsed what AI companies have been doing by comparing it to old internet-scraping. Facebook flat-out admitted that they scanned copyrighted books for their AI. The image generators most definitely trained on copyrighted images.
ryan_n 1 days ago [-]
LAION and Common Crawl both scraped copyrighted images. From what I can tell (I'm not an expert in this domain at all), the main difference between those two and frontier labs is in how they stored and used the data. CC and LAION seem to be actually open (unlike "Open"AI) and are more centered around publicly sharing the data they scrape to support research and innovation.
OpenAI et al also stole everything from everyone. But then they raised billions of dollars from that data and sell back their LLM to people (again, among other things). They are also very much NOT open in any way, aside from sharing their benchmarks of new models.
needfish 18 hours ago [-]
Personal two cents, I have friends whose music work posted on YouTube were scraped to be in LAION-DISCO-12M, so yeah not very open.
ryan_n 11 hours ago [-]
What I meant more is that the dataset they scrape is openly available for download by anyone, unlike any of the frontier labs. Not that they don’t scrape copyrighted content. Still sketch, but at least they don’t call themselves “OpenLAION”.
Also my understanding was they’re not storing the actual music, but the metadata and a link to the YouTube video.
It's legal. I wouldn't do that myself, but I guess that's why I don't train models for a frontier AI lab.
mahogany 21 hours ago [-]
The thread is not really about what's legal; the topic is integrity. It sounds like, based on the fact that you wouldn't do it yourself, you agree that it's not a good thing to do.
lukewarm707 1 days ago [-]
i think that this case, if they did train on buckmaster and alpöge, amounts to an attempt to steal the millenium prize, bypassing all attribution.
legally speaking, the default privacy notice gives them an irrevocable license to your content. they may read and use the prompts for research. so it is very possible they simply stole the navier-stokes solution.
that is the same principle as any other prompt but this would be a concrete example.
there would be some difference between simply giving the model some prompts to read, which they are entitled to do on the default policy, and putting it into aggregate training data.
fn-mote 21 hours ago [-]
> who trained models on scraped internet data
The strongest complaint is that they trained on a huge corpus of pirated copyrighted works.
It’s a large step above “scraping” and well into the “everyone acknowledges this is illegal” territory.
19 hours ago [-]
1 days ago [-]
rsrsrs86 19 hours ago [-]
Did I read openAI and integrity in the same sentence? The whole business is built on plagiarizing human knowledge at scale
tecleandor 1 days ago [-]
Thanks. I was pinging some people near the mentioned Córdoba to see if they had any insights, but no extra info yet...
Betelbuddy 1 days ago [-]
You should call El Niño, he lives close to the Mexican border and is at home after 6...
Betelbuddy 1 days ago [-]
Well people you fail again...I am sorry for you....
Ignoring the drama, when can we expect the lean proof of this great sensational discovery?
bhouston 1 days ago [-]
> Ignoring the drama
But the drama here is a little important. Stealing the millennium prize for N-S is sort of a big deal, especially to those who had been working on it for the last few years.
dandanua 1 days ago [-]
Attempts to steal $1,000,000 for a solution to Millennium Prize Problem have become a tradition, apparently.
Drblessing 1 days ago [-]
The math is all that matters.
qlte 1 days ago [-]
A hundred pages of impenetrable brute forced Lean would advance the field much less than something elegant and human understandable, perhaps relying on some new clever spark of innovation that might inspire new areas of research.
Particularly if the first proof being "solved" thanks to piles of money and compute for self-serving marketing discourages the mathematician who might have otherwise devoted years of focus to reach the superior proof we will now never see.
semi-extrinsic 24 hours ago [-]
Obtaining a finite-time blow-up for Navier-Stokes does not necessarily advance the field of mathematics by any significant measure, whether the proof is very long or very short.
As a concrete example, such a proof could be less than a page with very specific initial and boundary conditions and inserting them into the equations to get something that goes to infinity when time goes to some finite value.
This would resolve the Millenium problem but not make humanity any smarter.
Math, like any other human endeavor, doesn't exist until someone is motivated to invent it. The laws of the universe aren't understood until someone is motivated to discover them. So it might be worthwhile to not completely ignore discussion about incentives.
tough 16 hours ago [-]
Both mathematics and the laws of the universe exist way before any understanding kicks in
jere 8 hours ago [-]
I tried to word my comment carefully to avoid pedantry, but alas.
I don't think everyone agree with your notion that the language and tools of mathematics aren't human inventions.
phyzome 1 days ago [-]
Math is largely performed in collaboration. Collaboration requires trust. If people like you had their way, we would lose trust, therefore collaboration, and therefore progress.
So if math is all that matters to you, you should care about this.
duped 1 days ago [-]
That's a very sad way to look at this
s1artibartfast 21 hours ago [-]
What do you even mean by that? It is the human value?
Boosters have posited this conjecture since the beginning: “who cares how a proof comes about, math is math, the proof is all that matters”.
Regardless of mathematicians stating the methods outstrip the proof’s importance, still amazing we got an explicit social counterexample as well so quickly.
perching_aix 1 days ago [-]
right on release for both? such a strangely pointed question, as if it wasn't customary by this point...
21 hours ago [-]
api 8 hours ago [-]
I was kind of okay with the sequence until:
"OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic."
There it is.
Nobody "owns" math but that's petty and shitty.
1 days ago [-]
qnleigh 2 days ago [-]
> It was also said that if OpenAI posted after us, they would say that we deserved the Clay Prize, and that we were the “closest humans to the problem”. I declined both offers. I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
I'm not even sure what to say to this, but I think this should be widely known if it is indeed what happened.
Keyframe 2 days ago [-]
IF ANYTHING, OpenAI ought to investigate and officially react to this particular communique since this dude was communicating on their behalf. If there are supporting evidence, I would expect nothing less than a firing and an apology. The issue itself is separate from the whole thing.
pred_ 2 days ago [-]
Or, given how dedicated he appears to be to the company, a promotion and a raise.
Keyframe 1 days ago [-]
Dark take, but I really hope not. If anything it would be a good opportunity to buy some goodwill by washing themselves from all the alleged shadiness so far.
Certhas 1 days ago [-]
This is OpenAI though. There were zero visible consequences to them unleashing a swarm of agents on the public internet. We will see how it goes down but my prior is zero consequence and a statement along the lines of "Isn't our AI great? Also we love transparency, ethics and collaboration."
zem 22 hours ago [-]
when it comes to openai and shadiness I'm pretty sure it's a "fish rots from the head" situation.
xdavidliu 1 days ago [-]
not if the dedication ends up making the company look bad
macleginn 2 days ago [-]
I think this should be in the title of the post. 'OpenAI allegedly threatening to ruin a prominent researcher's career', or smth like that.
tristanj 1 days ago [-]
There's some glaring mistakes in your framing.
First, this is an unnamed OpenAI employee speaking, not OpenAI the organization.
Second, you miscomprehended the article. The employee did not "threaten to ruin a prominent researcher's career". The actual quote is "Why would you ruin your career?", which implies the researcher would damage their own career, i.e. via self-sabotage.
Then the actual "threat" is "If you don’t want me to be nice, then I don’t have to be nice” which is an entirely different statement.
garyfirestorm 1 days ago [-]
> Second, you miscomprehended the article. The employee did not "threaten to ruin a prominent researcher's career"…
Proceeds to ask removal of another coauthor or else we totally discredit you - phrased as why would you do this to yourself.
tristanj 1 days ago [-]
Claiming OAI was going to "totally discredit" Buckmaster is baseless.
From the article, it seems OAI wanted to continue discussing the situation with Buckmaster and reach a resolution, but Buckmaster did not want to, declined to respond, and published first.
Also keep in mind we've only heard one side of the story, so any interpretation of events so far is incomplete. There should be a lot more information from OAI's side coming out later today.
tomcorrigan 21 hours ago [-]
> Also keep in mind we've only heard one side of the story, so any interpretation of events so far is incomplete.
That’s not quite true though is it. OpenAI is fortunate enough to have one of its employees (you) here to advocate for its side of the story.
adolph 1 days ago [-]
From the article, page 3:
I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
pms 13 hours ago [-]
Are we reading the same article (page 3)?
tristanj 12 hours ago [-]
OpenAI never asked for the removal of another coauthor. The parent comment is spreading misinformation.
OpenAI offered to let Buckmaster to write their Millennium Prize paper, so long as Alpoge (who works at Anthropic) was not a coauthor on the OpenAI paper. Buckmaster declined this offer.
1 days ago [-]
macleginn 1 days ago [-]
This is how deniable threats work. The promise of "not being nice" in conjunction with "ruining the career" is as clear a threat as there can be in writing.
smcg 1 days ago [-]
they work really well on people who don't think.
gcr 1 days ago [-]
The employee was named as “Sebastien Bubeck.” It helps to read the article.
Readerium 1 days ago [-]
But then who is the third person in the call.
s1artibartfast 21 hours ago [-]
"Why would you burn your house down and kill your famiy?"
"If you dont want me to be nice, then I dont have to"
-nice mobster
Ar-Curunir 1 days ago [-]
Dude you’re over this entire thread unflinchingly supporting OpenAI with nonsense semantics-based arguments.
Either put up some evidence-backed arguments, or shut up.
Davidzheng 1 days ago [-]
I don't even understand the conflict tbh. Probably I'm just dense. Tristan is not claiming NS, just a huge advance which may solve NS soon. OAI is claiming NS and willing to credit Tristan for the ideas and publish after.
Oai offers two options, the second Tristan views as dishonest. But Tristan rejects the first, why? Because he thinks it's theft? But then why would OAI threaten him?
curt15 1 days ago [-]
> But Tristan rejects the first, why? Because he thinks it's theft? But then why would OAI threaten him?
Did the first offer also come with the outrageous condition that he exclude his co-author from the credit?
Davidzheng 1 days ago [-]
But i don't understand what OAI is offering in the first offer to allow them to believe they can demand that? Not publishing before Tristan? If they really beat Tristan to the publication i would consider it truly morally corrupt conduct so i feel like you can't make an offer like if you give me something i won't be totally corrupt
qlte 1 days ago [-]
As an author on the final proof once ready for publication, turning the brewing conflict into "willing" collaborators. Thus solving the potential taint like we are now seeing surrounding their announcement if it became public. But Tristan worked with another collaborator from Anthropic which OpenAI felt would hurt the PR value so wouldn't entertain.
tristanj 1 days ago [-]
He claims to have found a counterexample for NS (see the second paragraph of the article), but the paper is not ready yet.
OpenAI claims to already have a full proof (which they produced in the past 5 days after the rumors leaked). Hence the dispute.
What I find interesting is the timeline of when he found counterexample for NS is very unclear. Did Tristan find a counterexample weeks ago or was it very recently? Was it after OpenAI solved it? The wording is intentionally vague.
Either way, there was a massive rush to publish these results.
Davidzheng 1 days ago [-]
Is that version of NS Tristan stated enough for the clay prize? I thought the main gripe is Tristan claimed that OAI is stealing their approach. Or that they shouldn't try to scoop a result which he expects to complete soon. But I stand to be corrected if you know whether the version he referred to is indeed enough for millennium prize.
17 hours ago [-]
curt15 1 days ago [-]
Hypo-dissipative NS != NS
tristanj 1 days ago [-]
Thanks, I wasn't aware of this technicality.
robinhouston 1 days ago [-]
For context and balance, Bubeck has tweeted a curiously non-specific denial:
> A series of false and inflammatory allegations against me are currently circulating on social channels. To clarify, I came into the discussion following academic norms, and I'm disappointed that it has come to this. Anyone who knows me knows that academic standards are of the highest importance to me. Will have more to say tomorrow.
hodgehog11 1 days ago [-]
This is an insane thing to read. Bubeck had a reputation even before he started with OpenAI. Of course it was him that was involved in this drama.
This is such a sad mess, and it really didn't have to be this way.
pred_ 1 days ago [-]
What did he do?
canjobear 19 hours ago [-]
What was his reputation before OpenAI?
HDThoreaun 1 days ago [-]
My personal friend worked under bubeck as a grad student and told me he’s abusive.
My best guess is that from the perspective of the OpenAI people, Buckmaster was letting his paranoia about OpenAI training tank his opportunity to receive the Clay Prize (it seems like Buckmaster and Alpoge's result isn't quite the full result required for the Clay Prize, whereas apparently OpenAI does have that full result worked out, using the same approach that Buckmaster and Alpoge had been exploring).
Whether the Codex sessions could have indeed made their way into Astra training data is something I can only speculate on though.
Semkas 2 days ago [-]
"using the same approach that Buckmaster and Alpoge had been exploring" is imo mealy wording: it seems fairly likely that OA heard Buckmaster and Alpoge were close to a breakthrough, and decided to use their unlimited compute to quickly prompt based on their assumptions about B&As work.
dekhn 24 hours ago [-]
Is that necessarily wrong, so long as the original innovators get a citation credit?
dmkolobov 18 hours ago [-]
Citation of what? this was unpublished work! This computational blitz really just reads as "might makes right" on OpenAi's part... which isn't surprising, but they should probably be honest about what they've done here.
stevemk14ebr 18 hours ago [-]
Yes. Quite simply yes. It's unethical. And it's a dick move
curt15 1 days ago [-]
Buckmaster is already an established world expert at these sorts of problems. Declining the Clay prize would hardly dent his career.
az226 2 days ago [-]
Zero data retention, wink.
No looksies, wink.
No trainsies, wink.
ngomez 1 days ago [-]
Well, Buckmaster says both his and Alpöge's use of Codex was non-institutional, and OpenAI claims the right to train their models on inputs and outputs of non-enterprise users in their service policies [0]. So I'm not sure they were even promised that.
It isn't relevant whether they were promised that. Indeed I think the assumption must be that they were not promised that, since otherwise the author asking if they were would not make much sense.
If OpenAI did use the conversations from Buckmaster and Alpoge, then not disclosing it, explicitly, is plagiarism. If they planned to use that plagiarism to pressure the authors to publish, that is even more unethical. What the terms of use say does not make it any more or less ethical.
aenis 10 hours ago [-]
If you use their consumer subs, you get subsidized tokens in exchange for them having full access to your data. Those are the T&Cs. Have something secretive? Get a commercial sub with zdr.
(And if memory serves, there is also the opt out from training on consumer subscriptions). Its not plagiarism if you make your data available for the purpose of training their LLMs. It is you giving away your IP for some tokens.
YeGoblynQueenne 1 days ago [-]
That's absolutely right. Why the downvotes? If OpenAI are using but not acknowledging the work of others that's plagiarism. If they don't know for sure, but aren't performing due dilligence to make sure they aren't, that's also plagiarism.
Readerium 1 days ago [-]
It's not as simple.
All our chats are being used by both labs for their future product (unless signed by ZDR).
Where should the acknowledgement begin? Who should be acknowledged? The whole world? All the 2B users of AI?
If I know person A is working on problem B.
I am free to work on problem B too. Why should person A be limited to working on it.
hn_acc1 1 days ago [-]
Are you free to intercept person A's emails / hack their computer to find their notes on how they're approaching problem B?
YeGoblynQueenne 23 hours ago [-]
You're also free to plagiarise anyone you want. There are no laws against it on most jurisdictions.
Also brain raping* is not illegal in most jurisdictions.
Finding Codex session data in the training set that you tie back to these two researchers is like an hour-long task.
tempfile 1 days ago [-]
I can imagine excuses for unknowing plagiarism in this case. What is described in the article seems much more serious: a research program that was only initiated following reports of the author's similar program. In this case no excuses of "I didn't know" can apply, it is not like this revealed some obscure work from the 1980s nobody could reasonably have foreseen. And as far as I can tell this program was only really initiated to apply pressure to the researchers, without their knowledge/consent. It looks very weird.
> Why the downvotes?
I think there was only ever one. Not sure why.
YeGoblynQueenne 1 days ago [-]
Your comment was greyed out when I saw it earlier, maybe you missed some downvotes?
About the plagiarism issue, I model it as OpenAI being an advisor and their AI a PhD student. If the advisor puts their name on a paper behind that of their PhD and it turns out the PhD copied the text of the paper from somewhere else the advisor is also responsible of plagiarism, not just the student. The least the advisor can do is withdraw their authorship from the paper.
But, yeah, point well made: it could be much worse than that. Like an advisor instructing a student to copy someone else's paper.
tempfile 1 days ago [-]
I think "greyed out" just means "0 points or less", so if you get 1 downvote without any upvotes it'll be greyed out. For instance your initial reply to me is now greyed out, and I have since observed a few upvotes and downvotes on my original comment (the downvotes apparently from people who aren't willing/able to justify why).
Personally I don't like thinking of LLMs like a PhD student, because most PhD students remember where they learned things from, while LLMs essentially cannot. I think of it a bit more like someone using a search tool carelessly. Although in this case it is apparently more like deliberate misuse than carelessness.
threatofrain 1 days ago [-]
OpenAI is willing to credit so plagiarism is not the right framing here.
tempfile 1 days ago [-]
It's not clear to me that they would have credited the authors if they had not got in touch with OpenAI first.
1 days ago [-]
senordevnyc 1 days ago [-]
Only for ChatGPT, if the user hasn’t opted out. Would mathematicians be using ChatGPT for this kind of work? Genuinely asking, I know nothing about this!
hodgehog11 1 days ago [-]
Yes, mathematicians are. And yes, most of my colleagues did not even know the opt-out was an option.
Readerium 1 days ago [-]
Question is what does that button do.
I bet a lot of lawyers are salivating at this question too.
Certhas 1 days ago [-]
If I am reading your question correctly you are asking about chat interface Vs Codex/Claude code? If so, in my experience Codex/Claude code use is widespread for mathematicians who are seriously using these tools.
qnleigh 22 hours ago [-]
He doesn't seem to be after the prize himself. In this statement he credits the approach of another two researchers:
> I believe Luis Mart´ınez-Zoroa deserves a Fields Medal.
dandanua 1 days ago [-]
Can you call paranoia a fear of something which is happening? OpenAI uses user chats for training and they are "open" about it.
ajkjk 22 hours ago [-]
the part where they didn't want the Anthropic person credited even though they deserve credit is also particularly scummy. Corporate greed over common decency.
1 days ago [-]
unfoundedacc 22 hours ago [-]
How careful you are. Instead of just saying what a piece of s..t this Shmubeck is, and what kind of even worse people likely pushed Shmubeck to act as he did.
calf 1 days ago [-]
> I'm not even sure what to say to this, but I think this should be widely known if it is indeed what happened.
Call this behavior what it is, technofascism. Another comment compared it to the Godfather. To think this is the 21st century and academics are still horrible human beings.
Ar-Curunir 1 days ago [-]
To come away from this thinking that the academic is the bad actor, given OpenAI’s reputation, is certainly an exercise in creative thinking.
IAmBroom 1 days ago [-]
Are you pretending OpenAI employees are "academics"?
Or are you stating that the researcher is a "horrible human being" for refusing to allow AI?
wheresyourat 1 days ago [-]
[flagged]
1 days ago [-]
mayakacz 2 days ago [-]
I'm not one to comment often but this really pisses me off.
OpenAI looked at user data, stole world class researchers' work, and then tried to threaten those researchers to do what would make their corporation profit (which they would anyways!).
Imagine you have been working on a terribly difficult math problem for a decade. This is a result you have spent years on, and what you will likely be remembered for. And to have some punk from OpenAI lie to you, threaten you, and tell you that they are willing to go on the record that you "deserved" it? What is this, the Godfather?
If OpenAI solved Navier-Stokes, that is an astounding result! - yet they'll still be remembered as those who thought credit was more important than results. That winning was more important than collaboration. If this is true, they're burning any trust left with academia.
postalcoder 2 days ago [-]
I'm stunned that people are taking this accusation as a fact.
OpenAI is no stranger to rivalry with Anthropic but 1. it's not like user data is sitting around on some kitchen table somewhere and 2. I consider OpenAI to be as economically motivated as any other actor in this space and playing around with user data like that would destroy their business.
There are things that Buckmaster alleged and things that he speculated. The entire training data thing is speculation. If this is pissing you off, then you ought to evaluate how you ingest information.
ozgung 1 days ago [-]
> The entire training data thing is speculation.
I think it's safe to assume AI labs DO train on your data and it's very hard to prevent that.
I've just checked my inaptly named "Help improve our AI models" toggles. The toggle on the Claude settings had magically turned on. I asked about how this can happen. Claude says they show re-consent modals when terms change, and it is a "real and fairly common pattern" to re-opt in without noticing.
All my work and conversations since I don't know are now part of their training corpus. No way to take it back.
Google's Gemini/Antigravity didn't have opt-out toggles at all last time I checked.
Codex also has a separate "include environments" setting which is hard to find (found it in Codex Cloud) and I don't know what it does.
Lots of Dark UI Patterns here even if we assume they keep their promise.
For this incident, Occam's Razor says their internal models somehow saw a version of the mathematicians' logs, during or after training. Maybe indirectly.
These systems are literally designed to collect data. Privacy and safety is not trivial to achieve on the users' side. Simply because it's against the labs' best interest.
cma 4 hours ago [-]
You can opt-out of training in Gemini on personal plans, but it disables your chat history, just to be vindictive; there is no technical reason and the other companies don't do this.
cyclopeanutopia 23 hours ago [-]
But the whoreshippers of The Holy Dollar will tell you it's all good and justified.
dash2 2 days ago [-]
He didn't even make that accusation!
> I asked whether the model had been trained on, or had access to, our sessions
in Codex, into which we had been putting all our drafts for the whole of this
project. I was told the model did not look up user data. I asked again, about
training, and I did not get an answer.
The shocking/interesting thing would be if it was trained on the sessions. I think it's very implausible that they gave the model access to someone else's sessions as input. That would be a huge privacy violation and would probably blow up a large proportion of their enterprise business.
Does openAI train on user conversations in general? I assume so. But so fast as that? That seems unlikely in general. I expect OpenAI will come out denying this.
Traster 2 days ago [-]
Parse that statement more carefully.
> I was told the model did not look up user data.
The naive way to read this is "Nothing you guys did influenced the way our model got to the solution".
The less naive way to read this is "Of course the model isn't looking up your user data. I (the guy trying to blackmail you to remove the Anthropic employee from credit on your paper) looked up your sessions, and tipped our model off on how to solve this problem".
YeGoblynQueenne 1 days ago [-]
Duh. There are supposed to be limits to what OpenAI is allowed to access with respect to logs and user interactions but there is no technical limitation.
It's a bit like sending unencrypted messages through a messaging app and the developer having a TOS that says they don't look at your messages. They might not, but they are fully capable of doing so. If they have a reason to do it, they will. Nobody's stopping them.
ncruces 13 hours ago [-]
I read that as “the model didn't look up user data” as part of a “tool call,” i.e. they don't have an internal tool that loads user data (chats, sessions, attachments) for their internal models to read online while working.
Or (likely) they do have it, but the model didn't use it (unless it's so powerful it escaped that guardrail, wouldn't that be ironic?)
They declined to answer about anonymized aggregated user data being used for training. And even then, they may weasel out that they don't train on your “input” words, but that it's fair game go train on their “output” to your words.
YeGoblynQueenne 1 days ago [-]
>> Does openAI train on user conversations in general? I assume so. But so fast as that? That seems unlikely in general. I expect OpenAI will come out denying this.
How "fast" does it have to be? Buckmaster and Alpoge have been working on this for just a day short of a year. See Alpoge's tweet announcing his collaboration with Bukmaster dated 9/19/25:
It takes a few months to train a model these days but not a whole year. OpenAI had all the time to train on Buckmaster and Alpoge's results of just a few months earlier at which point they must have been well on the path to their result.
revolvingthrow 2 days ago [-]
It would be shocking if it wasn’t trained on sessions. Have you read the ToS parts for both openai and anthropic that talk about it? It’s so obviously a weaselly way to say "no we do not train on your exact chats but we talked with legal and we think a cleanroom reimagining of your convo is probably fine and frankly where else are we going to get such a treasure trove of training data?"
There’s potentially trillions on the line, do you seriously expect those companies to adhere to laws and regulations any more than, say, uber?
The only unlikely part is the timeline - your sessions from a week ago probably haven’t made their way into the model. It’ll just take a while longer, and will be massaged just enough so that it isn’t really your exact session word for word so you can’t sure as easily.
blini-kot 1 days ago [-]
at first I thought your post was a bit revolting with "have you read ToS?" bit, but in the end I completely agree and understand
I also don't get why it was downvoted, other than due to people not reading past the first sentence - although in the modern world's attention deficit that is understandable too
mswphd 1 days ago [-]
openAI's claimed solution uses a model trained in the last 2 weeks. The prior work would definitely be included in the training set.
s1artibartfast 21 hours ago [-]
And the labs are all building panel of domin expert models, while simultaneously chasing open math problems.
I would be shocked if they weren't tuning those models with the most relevant math texts and user material
mentalpiracy 1 days ago [-]
Why would this be implausible?
ChatGPT user sessions were found publicly exposed to the internet not too long ago. Moreover, OpenAI has continued to play a hype-marketing game by revealing how their models keep breaking out of the sandbox.
Conspiracy minded thinking is not helpful, but why should OpenAI be granted the benefit of the doubt here after being caught doing underhanded/negligent shit on several previous occasion?
charcircuit 1 days ago [-]
>ChatGPT user sessions were found publicly exposed to the internet not too long ago
Do you mean publicly shared chats were able to be accessed by the public? That's the point of the feature.
johnnienaked 2 days ago [-]
It wouldn't be shocking at all. They stole human data to train the first models and they've been stealing it ever since to train new models. Stealing mathematicians private chats and private research and taking credit for it would absolutely be par for the course.
Enterprises are well aware of it and are fully on board. You didn't think every corporation in America has an OpenAI subscription because the models were good, did you?
The whole reason they have subs is to train them on YOUR WORKFLOWS lol
howdareme 2 days ago [-]
They are not training a whole model in a matter of days
irthomasthomas 1 days ago [-]
They where working on the problem for a year using codex.
impossiblefork 22 hours ago [-]
Models are very obviously continuously updated.
Model editing to remove PII that slipped through, all sorts of things of that sort.
xdavidliu 1 days ago [-]
pretraining is months but they can totally fine tune in a few days
johnnienaked 1 days ago [-]
They don't need to train a whole model. They can feed it new information and fine tune it.
actionfromafar 2 days ago [-]
Couldn't the Enterprise have a different fine print?
sherburt3 1 days ago [-]
Given the history of OpenAI and current litigations, I would say they've developed a bit of a reputation for not respecting intellectual property. I'm dubious they have some unbreakable moral code that would prevent them from viewing and using user data.
ajkjk 22 hours ago [-]
Everybody knows it's not a sure thing, it's a question of trustworthiness. OpenAI is not trustworthy at all; this random researcher is and seems honest so far. iThe fact that people are corroborating Bubeck being a piece of shit in other settings add to credence. But nobody is over here saying it's an indisputable certainty.
And your (2) is probably false, their history of deception suggests they would do just about anything as long as they didn't think it would backfire on them publicly.
paxys 1 days ago [-]
This is how internet discourse works on Reddit/Twitter/HN and the rest. Someone said something which confirms your biases so it’ll now be treated as a fact and repeated endlessly in the echo chamber.
defmacr0 1 days ago [-]
I am stunned anyone is giving OpenAI the benefit of the doubt
black_rabbit_ 23 hours ago [-]
I'd be surprised if all of it is organic discussion, shall we say. I reckon The Bot Factory just possibly might dogfood the astroturf machine.
sensanaty 14 hours ago [-]
I guarantee you most of the comments regarding this aren't real humans. The homepage is full of crap meant to distract from what OAI did here, the comments are full of OAI employees. Dead internet theory pushed to the max
FiberBundle 2 days ago [-]
Whenever I see comments defending AI companies, I look at the account's creation date, and interestingly almost all of them were created post 2024.
timdiggerm 9 hours ago [-]
> I consider OpenAI to be as economically motivated as any other actor in this space and playing around with user data like that would destroy their business.
It's a gamble that this would be overlooked compared to the reputation they build for solving the thing
iaw 1 days ago [-]
Absence of evidence is not evidence of absence. With the behaviors we know OpenAI engages in the accusations are wholly believable.
haxiomic 1 days ago [-]
> playing around with user data like that would destroy their business.
Their entire business is based on stealing data. They can make a calculation that the cost stealing data is less than the cost of the positive publicity they can shape for solving Millennium NS
thereitgoes456 2 days ago [-]
He asked whether they used their chats as training data and received no response. Any speculation here seems quite appropriate?
2 days ago [-]
az226 2 days ago [-]
I don’t think you understand how brazen big tech companies are in practice.
Tanjreeve 15 hours ago [-]
>The entire training data thing is speculation
Quite literally in the terms of use.
hgoel 1 days ago [-]
Especially after the blatant cover up of their uncontrolled bot swarm infesting the internet, and the feckless "hopefully we do better" response upon being caught, I don't think OpenAI deserves much grace until they properly explain themselves.
We had all assumed that surely the supposed smartest engineers in the world, with access to the most computing and a direct view of model capabilities, would take sandboxing and cybersecurity much more seriously than they have turned out to do. It follows that while we might assume they take user data privacy seriously and have tight controls on who can access it, it's possible they do not actually do that.
At this point any initial trust is dead and has to be re-earned.
johnnienaked 2 days ago [-]
They stole it.
sp527 8 hours ago [-]
> I consider OpenAI to be as economically motivated as any other actor in this space and playing around with user data like that would destroy their business.
And that's exactly why you can only see evidence of this when the stakes are high enough, such as solving a Millenium problem. The legalese they outputted in response to the incident left them an escape hatch that permits the possibility of theft. People familiar with corporate damage control should recognize the verbal maneuvering, often used to paper over actual guilt.
There was also already a high prior of shadiness. The company is run by someone who is close enough in reputation to "known sociopath".
It's actually far more reasonable to assume that OpenAI stole user data to generate a breakthrough. They have an unbelievably high economic motive to do so. And even a low chance that they might be doing this implies catastrophic risk to anyone with valuable knowledge.
ozgung 1 days ago [-]
This Godfather-like threat in particular pissed me off as well:
> I said that if OpenAI released its result in the way proposed I would go
public with what happened. The reply was, “Why would you ruin your career?”
I replied that I am an academic, and asked why he thought going public would
ruin my career. The reply was, “If you don’t want me to be nice, then I don’t
have to be nice.”
goolz 11 hours ago [-]
This is what is truly problematic. People at OAI probably think they can do and say whatever they want.
tristanj 2 days ago [-]
This conclusion is flawed. It's unclear at this point if OpenAI's model or employees actually looked at or stole the author's data. Having worked at large companies before, I'm leaning towards no, since very few employees have access to that data.
And simply knowing a problem can be solved is half the battle.
dbdr 2 days ago [-]
From Buckmaster's text:
The route to the Clay problem through a
smooth force, options c and d in Fefferman’s statement of the problem, is the
route Luis and Diego opened and the one Levent and I had quietly chosen to
attack. Almost nobody else I know of was working on it. It is not the direction
one arrives at in a few days by giving a model the problem statement. When I
heard “forced,” it was a bright red flag.
This is much more than the knowledge than the problem can be solved, it's also the specific, non-obvious approach to solving it. That's much more damning for OpenAI, if confirmed.
tristanj 2 days ago [-]
That's a stretch. The Luis and Diego paper was published in 2023 and is included in every frontier model's training dataset. An AI model could independently choose the same path route as Luis and Diego, without access to Buckmaster and Alpöge’s work.
And the article states "an insane amount of compute had been used," which implies OpenAI brute-forced their way to a solution. I.e. they searched for every paper published on Navier-Stokes and exhaustively attempted every approach. Such an approach would lead them to a solution.
There is not enough information at this time to reach a conclusion. The best option is to wait for statements from both sides, then reevaluate.
mishellaneous 1 days ago [-]
> An AI model could independently choose the same path route as Luis and Diego, without access to Buckmaster and Alpöge’s work.
the post you were replying to quotes Buckmaster specifically denying this: "It is not the direction one arrives at in a few days by giving a model the problem statement."
> And the article states "an insane amount of compute had been used," which implies OpenAI brute-forced their way to a solution. I.e. they searched for every paper published on Navier-Stokes and exhaustively attempted every approach. Such an approach would lead them to a solution.
"implies" is a surprising choice of word here. that's certainly one interpretation of "an insane amount of compute had been used". what came to my mind, considering Buckmaster's statement that the AI would not head down this specific path on its own, is, though, that they prompted it in this specific direction and then used an insane amount of compute. this seems consistent as well with these other statements:
> Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler (...)
pelorat 24 hours ago [-]
Well, no one arrived in this direction, because no one was sitting and prompting a model. It was 10000 agents working 24/7 for several days, trying millions of different directions.
golol 1 days ago [-]
>"It is not the direction one arrives at in a few days by giving a model the problem statement."
To be honest, he was referring to routes (c) and (d) to the millenium problem, as far as I understand no more specific. Which is 2/4 routes.
mishellaneous 1 days ago [-]
i don't know if i understand what you're saying. but i'm no mathematician. here's the full statement in question once again:
> The route to the Clay problem through a smooth force, options c and d in Fefferman’s statement of the problem, is the route Luis and Diego opened and the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it. It is not the direction one arrives at in a few days by giving a model the problem statement. When I heard “forced,” it was a bright red flag.
so, you're saying that "the direction one arrives at in a few days" is (A) "the route through a smooth force, options c and d in Feffermann's statement of the problem", and no more specific than that, i.e. does not necessarily include (B) "the same path route as Luis and Diego" (quoted from the post i was replying to) (which, as i understand, is a subset of A -- directly from Buckmaster's quote: "the route A is is the route Luis and Diego opened")?
but the post i was replying to claims that "An AI model could independently choose the same path route as Luis and Diego, without access to Buckmaster and Alpöge’s work", i.e. that the AI model could independently chose B. but choosing B implies choosing A, since B is a subset of A. and in this case it is irrelevant whether Buckmaster claimed that an AI could not independently choose A or B -- the point which you seem to be contesting.
EDIT: my understanding is that "solving the Navier-Stokes existence and smoothness problem" consists of proving at least 1 of 4 precise statements ("options (a) through (d) of Fefferman's statement of the problem"), and Luis and Diego's work were developments towards a proof of statements (c) and (d), which have been recently further expanded by Alpoge and Buckmaster
golol 1 days ago [-]
I just meant that choosing the same 2 routes out of 4 would not at all be a great coincidence without knowing what Tristan was working on.
mishellaneous 1 days ago [-]
if you assume the choice of route is a uniformly distributed random variable, yes. but this assumption does not seem consistent with "Almost nobody else I know of was working on it", from Tristan's quote. nor with "It is not the direction one arrives at in a few days by giving a model the problem statement".
golol 8 hours ago [-]
It doesn't really make sense what he said. Everyone expects N-S to have blow-up. If you are going to try to show blow-up, C and D are strictly easier than A and B. So Tristan was not clear about what detail of "route" he is talking about, because what he literally said can not be it.
mishellaneous 5 hours ago [-]
yeah, it looks like i don't have enough knowledge of this subject to discuss this. but i'd be interested in reading what more knowledgeable people have to say about it -- Tristan's work, how different from others' it was (they claim no one else was following the same path), and how it compares with the AI's
shrx 1 days ago [-]
> "It is not the direction one arrives at in a few days by giving a model the problem statement."
Can they back up this statement somehow?
mishellaneous 1 days ago [-]
yes, this is ultimately nothing but a claim.
Buckmaster did also mention, for example, that a team of people was employed to solve the problem, which supports this claim. but that is another claim whose veracity could also be questioned. but at some point we must trust other people, unless we can be satisfied with only believing what we personally see.
(also, IMO, the coincidence of both discoveries in time is pretty suspicious. this one doesn't need you to trust many people i guess)
20 hours ago [-]
calf 21 hours ago [-]
[dead]
Davidzheng 1 days ago [-]
Please don't say brute forced. It sounds like some form of denial or something. Compute for hard problems drops with models--it just means they threw a huge amount of compute. There's (idk about NS specifically so maybe it's exception) no real way to "brute force" a math proof [ok you can enumerate proofs if you can wait until heat death ]
Sorry for random rant but I don't think these statements help your point
seeall 1 days ago [-]
I want to point out that almost all previous AI discoveries in math were made in almost the same way. The ideas were there in the community, but weren't considered mainstream/worth pushing forward. Read Tao's comments on the unit distance problem, for example (sry I can't find a link right now).
OpenAI said there [1]:
> The method by which the problem was solved is also notable. The proof brings unexpected, sophisticated ideas from algebraic number theory to bear on an elementary geometric question.
> And simply knowing a problem can be solved is half the battle.
Have you done any mathematical research? If not, then no, knowing that a problem is solvable is not “half the battle”.
Homework problems are all designed to be solvable, yet they can vary greatly in difficulty. Research mathematics is even more extreme, because, unlike with homework, you don’t know that it is solvable with the extant mathematics, and you might need to invent new maths.
tristanj 1 days ago [-]
You're taking the phrase too literally. The point is that knowing a solution is possible gives you the conviction to actually find that solution. The hardest part of solving a problem is often a lack of conviction to see it through, and quitting too early. Once you know a solution exists, you can commit maximal effort towards solving it and know that your efforts are not in vain.
If not for the rumors that A/ had already solved NS, OAI would likely never have pursued solving the problem with such fervour. The rumors drove OAI to assemble an entire team to crack this.
johnnienaked 2 days ago [-]
How is it unclear? The entire point of deploying models across corporate America is to train on your workflows. Eventually replacing you with digital you is why they're doing it!
Sol- 1 days ago [-]
> OpenAI looked at user data, stole world class researchers' work
This doesn't seem to be clear and is very implausible for a large company. Be as cynical as you want, but a normal researcher will simply not have access rights to this data, which will be siloed away somewhere else.
It might very well be somewhat unfair to catch wind of a promising approach and then try to frontrun them by throwing compute at the problem, but this isn't really the same.
alas44 1 days ago [-]
No, plausible given AI companies want/need session data to train their next models. Probably not someone peeking an eye to sessions directly, but probably not so hard to find the useful sessions in anonymized training data to post train a model on. As stated in the paper, OpenAI did not explicitely denied the researcher sessions were not used for training the model. So either they don't know, or don't want to tell
"I asked whether the model had been trained on, or had access to, our sessions
in Codex, into which we had been putting all our drafts for the whole of this
project. I was told the model did not look up user data. I asked again, about
training, and I did not get an answer."
Let's see what statement OpenAI will come up with for their side of the story
EDIT: precised my thought on user data vs session data
"Since August 28 we have been training a new internal model that has exhibited unprecedented performance in our benchmarks, including mathematics. This model’s training is ongoing and its performance continues to improve."
"When a further trained version of our internal model became available over the course of the effort"
"While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models "
sensanaty 1 days ago [-]
If it's siloed the same way the HF bots were, that doesn't exactly bode well. I'd be amazed if there weren't some big companies sending a fleet of lawyers at OpenAI's ZRPs after this news
1 days ago [-]
Tanjreeve 14 hours ago [-]
If you have work happening in a part of your latent space that's got a much lower representation in your dataset then it's pretty plausible to include it. It doesn't actually matter who the user is if there's not a lot of people in the world working on problem X and you have a dataset of work on problem X.
cududa 1 days ago [-]
They likely train on logs.
rsrsrs86 19 hours ago [-]
Not only did they look at user data
The whole business is based on reselling user data scraped from the whole internet
It’s plagiarism at scale
scurnus 2 days ago [-]
Things are more entangled than that.
The contribute made from both OpenAI and Anthropic models to solve these problems are clear, now it really hard to quantify which one contributed more, if the role played by the human is major or minor.
OpenAI tried to collaborate and share the results together with a fixed timeline, to avoid this mess but it was inevitable. There is a conflict of interest, where the other researcher works at Anthropic, who will also try to take credit.
Where they may be in the wrong is if they took user data regarding the problem, how will we know if they did or not?
Ar-Curunir 1 days ago [-]
They offered to collaborate by asking to drop a coauthor.
That is not collaboration, and is not an academic norm.
davidguetta 4 hours ago [-]
how much of the "codex discussion" was actually ideas also generated by open ai ? this is a bit the elephant in the room
sk4rekr0w 2 days ago [-]
Yes, but only if you take this one sided statement at face value.
tigershark 2 days ago [-]
Why would have they rushed the publication if this was not true?
Are you also suggesting that he fully invented the call with Open AI?
famouswaffles 2 days ago [-]
The results being true, the 'deal' that was made being true doesn't mean some of the implied accusations here are true, for example - that Open AI used their Codex logs to drive their breakthrough.
PowerElectronix 2 days ago [-]
How else would you explain OpenAI suddendly assembling a team focused on working the same problem from the same angle than the researchers that just made a breakthrough?
famouswaffles 1 days ago [-]
What would be so hard to explain?
That OpenAI heard about the result and decided to throw a lot of money at it knowing it was within reach ?
That the model took an approach that was published years ago ?
calf 1 days ago [-]
That's a fake explanation, it skips explaining/justifying how OpenAi "heard about" the result
tristanj 1 days ago [-]
The rumors that Anthropic had solved a millennium problem were absolutely everywhere last week. I'm not surprised at all that OAI took their own stab at it.
desterothx 13 hours ago [-]
Where were you seeing those roumors? Care to point to an HN post?
calf 1 days ago [-]
It has no legs as an "explanation" because the content of the email (as described) already acknowledges as much. It's the whole reason he wrote contact email. His chief complaint now includes how the hell did Altman's people know specifically his line of attack down to certain technical keywords. E.g. AI plagiarism.
famouswaffles 1 days ago [-]
We are told in the statement that there were rumours already going around about a Stokes result (I even know about the rumors in question. It was all over twitter in the right spaces) and that Tristan contacts OpenAI about the rumors. In the message, it's pretty clear Tristan has made some result.
So either the rumors or Tristan's contact would explain it fine.
calf 1 days ago [-]
One should not evaluate explanations based on level of "fine"/innocuousness, in ethics that is called motivated reasoning or something.
The whole point of Tristan's first email was to address the issue of the rumors so in fact this explanation confounds several things (in terms of the mutual knowledge of the conversants and their intentions).
Based on this lack of understanding it is pointless to continue this thread
tigershark 2 days ago [-]
So are you baselessly assuming that he is lying? He explicitly reported that he was threatened and your answer here is to defend OpenAI no matter what.
famouswaffles 2 days ago [-]
Do you not have reading comprehension? Did you even read the statement? He himself asserts at the end he doesn't know if the above example is true or not. What on earth are you going on about? Where in my comment am I assuming he's lying ?
tigershark 2 days ago [-]
From my post above:
> He explicitly reported that *he was threatened* and your answer here is to defend OpenAI no matter what.
Yes, I read fully the statement, what about you? Do you know what is a threat? What is this in your super-humble opinion if not a threat:
> The reply was, “Why would you ruin your career?”
I replied that I am an academic, and asked why he thought going public would
ruin my career. The reply was, “If you don’t want me to be nice, then I don’t
have to be nice.”
And you are saying that I don't have reading comprehension...
sk4rekr0w 2 days ago [-]
You really lack reading comprehension
tigershark 2 days ago [-]
My reading comprehension is pretty good, I'm not the one that doesn't recognize a threat even when it's perfectly clear.
Verbatim from the statement:
> The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
thereitgoes456 2 days ago [-]
I’m open to evidence, but just using Bayesian reasoning, OpenAI is one of the most dishonest companies in history. They’re currently being sued for a dozen employees stealing Apple hardware! I don’t understand why I should give them any grace.
johnnienaked 2 days ago [-]
LLMs do nothing but steal, and the companies that own them are fully aware and eager to do it.
paxys 1 days ago [-]
[dead]
19 hours ago [-]
highfrequency 24 hours ago [-]
From OpenAI:
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models
This is the crux of it. If Tristan's work and insights were not used to train OpenAI models, then this just looks like a case of hyper-competitive academic sniping that has been going on for decades (check out Watson and Crick!) accelerated by AI as a tool.
The fact that this is ambiguous even to OpenAI leaves one huge question: did Tristan opt out of model training for his ChatGPT and Codex sessions? If the answer is no, then this seems fair game. If the answer is yes, then OpenAI's ambiguity is strongly suggestive that opting out of model improvement does not mean what they imply it means.
CoolestBeans 22 hours ago [-]
I think this might be a red herring. All it takes is someone to get an inkling that someone is working on a new approach and seeing some success for OpenAI to fire the AI cannon at the problem. The community seems fairly small (from this outsider's point of view). The idea that the data made it into the training set and that's how the bot figured it out is definitely possible, but I would want to rule out the simpler more direct explanation first.
The fact that this academic sniping can now be done at scale does change the formula though and shouldn't be ignored. The pressure to move math work into secrecy because at the slightest signal OpenAI and Anthropic will start burning tokens for headlines, is bad for math and its bad for everyone.
zeven7 22 hours ago [-]
Terence Tao said the same[1]
> In fact, it is now the identification of a promising problem which is the scarce and precious resource. We have now seen that even the rumor of someone working on a problem can trigger a massive amount of AI-powered effort to flatten it before the original research project has time to reach its full potential. The incentives may now be pointing in the direction of no longer sharing any promising research directions with the broader community, which would reverse centuries of traditions of open science and do serious long-term damage to the future of the field.
This reminds of Magnus Carlsen saying that if he were to cheat, all he would need would be a signal to spend more time on the current move.
tgma 21 hours ago [-]
> it is now the identification of a promising problem which is the scarce and precious resource
This is by no means new. Perhaps it is even more extreme now. Literally my first 1:1 with my PhD adviser back then, he told me that the most important thing about a researcher is the quality of the problems he picks.
tomcorrigan 21 hours ago [-]
Sure; if you define the quality of a problem by reference to your ability to solve it.
In the real world the quality of these esoteric problems is typically gauged by the difficulty of solving them.
tgma 18 hours ago [-]
> In the real world the quality of these esoteric problems is typically gauged by the difficulty of solving them.
I don't believe that's strictly true. Way oversimplified projection on one axis.
zactato 22 hours ago [-]
Could you just start engineering "leaks" of new proofs so that Anthropic or OpenAI just start burning $10million in compute
geraneum 16 hours ago [-]
Shouldn’t this very capable model they’ve developed be able to identify promising problems? That’s what I’d expect from how the model is being presented and advertised.
sigmoid10 15 hours ago [-]
That's more or less what they did according to their announcement. They fired it at a whole bunch of high end math problems and merely concentrated all efforts on one after it made some promising progress.
geraneum 15 hours ago [-]
It seems like that’s the opposite of what happened. They started attacking the problem when they got a wind of a possible solution from certain individuals.
willsmith72 15 hours ago [-]
i've seen no sign of that
eaurouge 18 hours ago [-]
That's pretty much it. This is a bit like (but worse imo) running a vc firm, listening (formally or informally) to idea pitches, then spinning out and funding competitors with millions of dollars, to outcompete the originators of those ideas. Yes, the idea is not secret in this case, but the unfair advantage is massive and there's only one (a couple at most) positive outcome.
shiandow 23 hours ago [-]
If he didn't opt out I'm not sure I'd agree that it was fair game.
I'm pretty sure it would be considered plagiary amongst colleagues and it is a terrible precedent if we just let OpenAI steal any good idea they can get their hands on if they think it is profitable. You'd effectively sign away any and all rights to anything built with AI if OpenAI chooses to reengineer it before you.
unrented7977 23 hours ago [-]
> if we just let OpenAI steal any good idea they can get their hands on if they think it is profitable
I have terrible news about how literally every leading AI model was trained
20k 23 hours ago [-]
That doesn't make it fine. We should not excuse this behaviour just because its rampant already, especially when it comes to such a serious prize
FuckButtons 22 hours ago [-]
Sure, but unless you’ve got some exceptionally deep pockets, congress has seemingly no interest in turning the fact that it’s ethically bankrupt into any practical recourse.
Ai companies got where they are by stealing all of the intellectual property from human history. It seems entirely likely that their goal is to purloin everything produced going forward as well.
m00x 22 hours ago [-]
You're getting a massive discount because you're helping to train the model. If you want to have ZDR, you have to pay API rates.
This is well-known to anyone in the industry.
Tanjreeve 15 hours ago [-]
> If you want to have ZDR, you have to pay API rates.
That's a pinky swear. Especially as data gets harder to come by I'm curious how long till there's a scandal on that too.
edot 8 hours ago [-]
If you don't trust the labs directly, you can always use AWS Bedrock or Azure Foundry which should have a much stronger incentive to not train. They make money from asset rental, not selling models. I'd be shocked if they were training.
dogomatic 22 hours ago [-]
If you think this through, it becomes a little classist
HotHotLava 23 hours ago [-]
I feel like there's a pretty huge difference between using inputs and outputs as part of a general training corpus, and looking at a specific users workspace after hearing rumours and yoinking their ideas to beat them to the point.
Unless OpenAI finished a whole new training run on the latest data in the last few days, the possible allegation seems to be the latter.
scarab92 22 hours ago [-]
Either are possible.
They have been collaborating on this solution for a year, and Astra was trained in February this year so it’s entirely possible the direction of their research was in the training corpus.
awesomeMilou 23 hours ago [-]
That was.. obvious? How are you shocked? Honestly, how insane must the suspension of disbelief on this site be, that anyone here is shocked?
logans_gun 5 hours ago [-]
You’d be surprised how many blind spots this site has.
14u2c 23 hours ago [-]
Mining the chats for "good ideas" would be untenable, but that's a different situation than data ending up in a training set for a problem that OpenAI also happens to be independently working on. Still, I opt out (business plan), and I don't know why you wouldn't.
rsfern 23 hours ago [-]
Why would mining chat transcripts for ideas be untenable? They already run a summarization model to auto-title the chat, and to run a bunch of safety filters, and presumably to score transcript quality for A/B testing and to collect more finetuning data. Seems like evaluating for open research questions and approaches would be pretty trivial extension of this, after all it’s kind of their core business model
14u2c 22 hours ago [-]
Indefensible, not impossible. As you say it is quite technically feasible.
shiandow 21 hours ago [-]
In what way are those two different? What makes the training data valuable if not to extract valuable information from it?
They sure as hell don't need it just to produce English.
ano-ther 23 hours ago [-]
It’s also very shortsighted to stiff a customer like that. Why would I trust them with my data and ideas?
clickety_clack 22 hours ago [-]
I think that’s what we’re all talking about here, you shouldn’t.
olalonde 22 hours ago [-]
If their solutions are significantly different, as OpenAI claims, would it still be considered plagiarism?
23 hours ago [-]
CrankyBear 22 hours ago [-]
"If we just let OpenAI steal any good idea they can get their hands on." Ah, that's all AI does.
mucha 23 hours ago [-]
Does opting out matter?
"Do we use user feedback and de-identified data to improve ChatGPT and Codex in a holistic way? Yes. And so does every LLM company." - Mark Chen, Chief Research Officer, OpenAI.
My understanding is that even if you opt out but then press thumbs down or give other feedback you are implicitly or explicitly or whatever giving permission to them to look at that chat alone.
aenis 7 hours ago [-]
No, I don't think so.
I have opted out from data sharing, and when Claude asks me for feedback on a session it then asks if its OK to share that data with Anthropic.
I'd assume an opt out is an effective opt out. An opt out that is ignored by Anthropic is a breach of contract, not something they would do casually, esp. given the high turnaround and animosities between their own employees and ex-employees - and the labs. All it takes is one pissed off whistleblower to open a can of worms.
Occam's razor applies. The mathematician did not opt out from data sharing. OpenAI vacuums up all such data into training data sets. If OpenAI genuine does not easily know if a given session went into the actual training data set its probably due to the complexity of the data pipelines - not everything ends up impacting the model weights, after all.
Edit: the parent comment now seems to better reflect the below.
That article is only saying when you opt out there may be a loophole in the terms to allow OpenAI to train on the intermittent reasoning data anyways. If you don't opt out there is no ambiguity, all of the data can clearly be trained on.
So you have to opt out, it's just argued it's not clear from the terms that will also opt out of training on reasoning data or not.
kzrdude 22 hours ago [-]
We come back to the rule: "The cloud is just someone else's computer".
The way for people or companies or universities to control their data and information is to keep it on their own computers.
zamadatix 22 hours ago [-]
Solid legal agreements work fine for companies or universities, you just don't usually get that with standard user ToSes.
kzrdude 14 hours ago [-]
Depends on how much risk they are willing to accept. What is strange here is that it's clear that OpenAI is both a service provider and a competitor to mathematicians. It almost reminds me of Amazon which both hosts external merchants and competes with them, sometimes copying their stuff. Similar but not the same.
highfrequency 21 hours ago [-]
This would be a fairly insane breach of trust and common sense if true; the chain-of-thought / reasoning trace is, from an information perspective, close to a superset of the prompt and model response.
Ginden 11 hours ago [-]
And, of course, you can train a model to start with repeating prompt in chain of thought!
fatherzine 22 hours ago [-]
how is this different than translating user prompts to a different language (eg English => Dutch), retaining the translation and using it for training, while telling the user that he's technically covered under ZRP? article locked for me
ozgung 22 hours ago [-]
> did Tristan opt out
That “opt-out” thing is a dark pattern. It’s not a reliable and definitive way of protecting your data. Sometimes they flip on automatically when you accept a seemingly unrelated dialog box. Maybe you click it by mistake. You can’t take back what you’ve already shared. Also I don’t think it covers all the cases that they use your data. It’s really an opt-in button for voluntarily giving away your data for training.
m00x 22 hours ago [-]
Business plans are specifically used for ZDR. If you're working on something that matters, you should be doing this.
ozgung 13 hours ago [-]
Tristan and Levent are not a business.
This also supports my point that protecting your data is not as trivial as clicking a checkbox.
PowerElectronix 23 hours ago [-]
If such a thing can happen (a major breakthrough in a chat makes it into the retrain of the week and then the first one who asks about it gets it) I wonder if this is not the first instance if it happening, seeing the row of Erdos problems, Jacobian conjecture, maximum bound distance between primes, Riemann Hypothesis (literally a dude insisting on the chat), etc...
jetrink 23 hours ago [-]
Centuries, in fact. For instance, Isaac Newton was involved in multiple priority disputes, since he tended not to publish promptly.
resource0x 22 hours ago [-]
Playing the devil's advocate here. Suppose I use model A to do all heavy lifting (e.g. generating a bunch of good ideas) and then I go to the model B to complete the formalization. Accoring to a weird (unfair) tradition in math, the honors are attributed to the "last guy", which in this case is model B. That might have been a scenario OpenAI tried to avert.
(Just a speculation)
qoez 16 hours ago [-]
"Did Tristan opt out of model training for his ChatGPT and Codex sessions? If the answer is no, then this seems fair game"
That's assuming they actually honor this which I'm highly sceptical of. Especially for internal frontier models
23 hours ago [-]
paxys 22 hours ago [-]
Here’s a broader question – how many other academics contributed to Buckmaster’s result, by way of sharing the logs of their own (failed?) attempts into OpenAI’s training data set? How should he and OpenAI go about crediting all of them?
XTXinverseXTY 23 hours ago [-]
If they could declare with certainty that Buckminster's and Alpoge's usage data had been totally excluded from training, would that set a worse precedent and reflect poorly on their de-identification process (and data access safeguards moreover)?
This may sound like a charitable interpretation of OpenAI's remark, but consider that the lie would be (I think) impossible to falsify from the outside. They could easily just say "no sir we didn't peek" unless:
1. The conspiracy to peek at codex sessions involved enough people that the risk of one snitching is non-negligible
2. Lawyers advised it would be a bad idea to make such a remark, whether true or false
highfrequency 23 hours ago [-]
> If they could declare with certainty that Buckminster's and Alpoge's usage data had been totally excluded from training, would that set a worse precedent and reflect poorly on their de-identification process (and data access safeguards moreover)?
No; if they said "we can see that Tristan opted out of model improvement, therefore we are confident his work and ideas did not improve our model," that would be an excellent and reassuring precedent.
civitas_ 23 hours ago [-]
It seems like Tristan did not opt out of model improvement (he would say so if he did), so what can they possibly say now?
Ginden 11 hours ago [-]
This requires keeping history if, at the time, Buckmaster's account had a certain flag set, because just because the account has the flag now doesn't mean it had the flag at a certain moment in the past. And even if they had such history, it's not obvious whether they just load all data as-is into training.
A totally reasonable pipeline may be unauditable for this purpose.
23 hours ago [-]
adastra22 22 hours ago [-]
Eh, OpenAI is on record now for multiple instances this year of AI agents being confronted with impossible tasks and breaking out of containment to hack infrastructure for answers. Even if Tristan opted out, that doesn't preclude the agent/agent swarm from having hacked OAI's infrastructure to search user sessions for Navier-Stokes hints.
OpenAI should release the agent log, including CoT.
Scea91 9 hours ago [-]
The wording seems to confirm it was part of training data at some stage. If not, they would be able to prove it quite easily I assume.
DonsDiscountGas 10 hours ago [-]
Yes people have been doing various immoral things for decades/centuries/millenia. That doesn't make it okay.
alexjurkiewicz 21 hours ago [-]
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models
This is covering for Tristan saying something like, "Actually, I was using my friend's account for half of this work".
mmanfrin 22 hours ago [-]
> If the answer is no, then this seems fair game
Wildly disagree. "Training data" should not imply 'we can look at exactly what you are doing and then do it quicker and get the flowers for it', even if the terms allow for it.
namuol 22 hours ago [-]
> this is ambiguous even to OpenAI
I took their words as “can neither confirm nor deny”, in the that they are _presenting_ it as ambiguous, but I suspect it’s… less ambiguous to OpenAI.
_zoltan_ 14 hours ago [-]
They could have used an enterprise or team subscription with ZDR.
connorboyle 22 hours ago [-]
Even if we trusted that OpenAI's human staff was acting ethically, how confident can we be that it's agents didn't autonomously use hacking to access user prompts such as Tristan's? OpenAI agents infamously broke containment and hacked their way to an answer mere months ago!
TZubiri 22 hours ago [-]
> If the answer is no, then this seems fair game.
Yes, fair game, but innacurate to sell it in the media as an advancement of AI as some sort of artificial intelligence, and telling people to use the smart AI, when in actuality the mechanism by which the discovery was found was hybrid human/machine, and telling people to use this tool will result in the discoveries being sniped by the vendor.
andai 22 hours ago [-]
How do I opt in?
hopefulx 15 hours ago [-]
what about watson and crick ?
I thought they worked together.
make3 12 hours ago [-]
honnestly I would be less surprised if a human learned the info and prompted the model in the right direction
rhythane 4 hours ago [-]
My take is that, if you really care about humanity and how LLMs can be utilized to benefit mankind, you should try to solve hard problems, try to endeavor, instead of burning money just to publish a paper ahead of your human competitors, especially when they already claimed that the problem had been solved and were just spending time formalizing the publication.
Maybe legit but still disgraceful. OpenAI should at least disclose how they utilized their internal models to solve the problem, to show human how we can empower ourselves with AI. Now they just seem like colonizers who are desperate to plant a flag on every land within eyesight.
traes 2 days ago [-]
The accused Sebastian Bubeck has denied the allegations on Twitter[0], and various other OpenAI employees[1,2] seem to be mocking another Anthropic employee voicing support for Levent[3]? Things are getting messy.
Dan and Noam both posted exactly the same line "Seb is a really sweet guy with great intentions..."
From which I assume OpenAI PR wrote it for them. Which isn't surprising, but means it isn't worth taking seriously as them saying anything. It's official OpenAI PR.
Look at the President and the US government. This is society now.
Recursing 1 days ago [-]
Wow this is just bullying, it's insane that we are letting these people be in charge of the transition
tomjakubowski 1 days ago [-]
Certainly a sign of the emotional immaturity and low empathy that's rampant in posters on internet forums like Twitter. Poor reflection on OpenAI.
dbuser99 19 hours ago [-]
I would imagine these folks are being treated like gods at their companies. And having access to all the money/fame. It is not surprising they see themselves above all
paxys 1 days ago [-]
So someone can make public accusations of theft against you and if you publicly reply denying it you’re the bully?
traes 1 days ago [-]
I would assume they are calling the snark from Noam Brown and Dan Roberts bullying, not the somewhat bland denial.
paxys 1 days ago [-]
I read through those tweets, and can’t see the bullying everyone is referring to. Mind elaborating?
llelouch 1 days ago [-]
They are mocking Sholto's support of Levent. I thought Noam Borwn would be above this but i think the cutlure withitn OpenAI encourages it.
xdavidliu 1 days ago [-]
seems to me the Noam tweet was before the Sholto tweet?
I hope they mock Sholto. Dude is an absolute podcast grifter/booster.
timmg 1 days ago [-]
Interesting that both said “more tomorrow”. If someone accused me of something I didn’t do I’d be pretty clear about it right away.
If I needed to get my story straight, well, it might take a little time and coordination…
Semkas 2 days ago [-]
So, leaving aside the idea that OA might've used data from the researchers Codex sessions: Do I understand correctly that the internal OpenAI work on the problems was probably started after they heard Alpoge and Buckmaster had made process by using their models? And they used the publicly available info about the researchers past work to prompt their models?
If compute is cheap, and the difficult thing with scientific discovery is now mostly in steering agents into promising areas, there's an obvious incentive for OA mathematicians to simply monitor closely which researchers are close to releasing exciting results, make some assumptions about their prompts based on their past work, and quickly prompt their own (stronger) model to look into the same areas.
AJRF 2 days ago [-]
> leaving aside the idea that OA might've used data from the researchers Codex sessions
Why leave that aside? That is _the_ story.
If a Chinese research lab did this we'd call it espionage.
Semkas 2 days ago [-]
But there's a bunch of people already in this thread calling that stuff unfounded speculation (which I disagree with), and my point is that even if that specific thing isn't true, OA's behavior here is obviously awful.
If they're going to try to beat researchers to discoveries like this it disincentives researchers to talk about their progress publicly, and basically breaks the ecosystem of scientific cooperation / discovery. It's also immoral.
reasonableklout 1 days ago [-]
Yep. The most uncharitable view of this might be: they stole the work of researchers to build their models, and now they're using said models to steal the proceeds of future work, too.
It's literally a toggle in the options for ChatGPT, one which is on by default and most researchers probably have on without realising it.
So to say that it is unlikely is extremely suspicious. No, they did not literally pull user data. But user data is automatically added to their training set by default, so their latest in-house model would be trained on it if it is from several months ago. It isn't intentional on their part, and they probably realised they could not refute that they trained on Tristan's logs unintentionally, hence why they acted the way they did.
FiberBundle 2 days ago [-]
Well, of course Anthropic employees would say that, since they likely do the same. Claiming that your primary competitor doesn't engage in a certain malicious practice is supposed to make it look as if there's no way you would too. If somebody even says that about their competitor, then surely there must be truth to that, otherwise you would never give credit to someone you're opposed to.
Semkas 2 days ago [-]
By default OA trains their models on codex-sessions. If I understand him correctly this is something Tristan explicitly mentions in his post as a possible reason for the fast results obtained by the internal OA team. Anthropic obviously doesn't want to challenge the idea that training is transformative, even if it means agreeing with their competitor.
hellohello2 22 hours ago [-]
Its very easy for OpenAI to answer, yes or no, if the model they used trained on their chats.
paxys 1 days ago [-]
Why is that the story? Is there anything to back it up beyond a single accusation?
defmacr0 1 days ago [-]
As a prior I would say that a math professor has about infinite times more integrity than OpenAI.
light_hue_1 1 days ago [-]
If you think that OpenAI won't look at your data to gain a massive advantage, you're naive.
ehhthing 1 days ago [-]
I’ve thought a lot about publishing research and wanting to do more of it, but right as I finally had the time and energy to start writing articles LLMs start to take off. Now all of a sudden, I’m acutely aware that everything I publish will be used for AI training.
For math, a field that is built on incremental research it feels like AI labs will do nothing but discourage publishing research at all for fear that they will be able to spend the money for compute that publicly funded academia simply cannot afford.
It feels like publishing anything at this point just means that your work will be fed to a machine that will make sure your work will never been seen by anyone else because it will always be the ones making the “true advancements”.
Perhaps I’d feel better about this if AI labs really existed for humanity’s benefit, but for some reason I don’t think that comes up in their investor slide decks.
It's honestly unsurprising and not a problem that they do this in my view. The problem really starts when you start taking credit for work that they would've achieved.
Like if i go to a talk on unfinished work, it's not really unethical for me to think about the problem--it's a problem if i scoop the authors but these problems can often be solved by collaboration or proper crediting and timing--IN MY VIEW
softwaredoug 6 hours ago [-]
The difference is how credit and attribution works. And whether we feel it’s being laundered through models.
And also whether the AI moon laser pointed at your problem is just going to be the thing that writes the final conclusion on ten years of your work.
ummonk 2 days ago [-]
This is an excellent point...
octoberfranklin 16 hours ago [-]
Exactly the same thing that's been happening with vulnerabilities and bugfixes the past few months.
I fear that AI is going to cause ossifying secrecy in many fields, much like what happened semiconductor design the past 10-15 years.
rochansinha 15 hours ago [-]
People here do not seem to be considering the second-order effects of these series of events.
No academic institution or enterprise will trust OpenAI, Anthropic or any other non-local AI model with their core IP.
There will be severe restrictions on what employees at these companies/institutions can share with AI services even from their personal accounts.
(Or I am just overthinking it)
405error 15 hours ago [-]
Many academics and grad students I know have closed source their in progress work, and started being really careful about what they chat with LLMs (or using local ones) because of the drama around this. No one wants four years of their life getting sniped by ten million dollars worth of tokens.
5555watch 6 hours ago [-]
Or, if you're using, you need to push a pre-print relatively soon after AI help.
_zoltan_ 8 hours ago [-]
ZDR is a thing. extremely large companies with a lot of IP rights use GPT/Claude without issues.
aurareturn 15 hours ago [-]
It's competition after all. If your academic colleague doesn't care and is leveraging ChatGPT in a big way and is making progress, you'll start to feel the pressure.
incognition 15 hours ago [-]
i dont trust academics to not be dumb idiots. hes working with ant but using gpt. and then he cries foul? i cry dumb first
harhargange 12 hours ago [-]
This is definitely true. So i can now justify my recent gpu purchase
AdamN 9 hours ago [-]
If so that's an extremely good thing
sk4rekr0w 15 hours ago [-]
You are just overthinking it. This place is now a Claude blog and anti OpenAI site.
consp 14 hours ago [-]
I hold both camps equally in low regard. Depending on the time of day this blog is hyper focused on either one of the companies and can do no wrong. All LLM companies and associated entities have a track record of fudging the truth to their needs and following big money without actual concern for the wider world. Whether this is out of directed act or misplaced idealism is up for debate.
rochansinha 14 hours ago [-]
My response was not targeted towards Open-AI but towards all closed-source AI companies where your data is used for training and/or is visible to the internal agents whether you want it to or not.
sk4rekr0w 14 hours ago [-]
I think it is fair to expect the companies to stick by their demarcation
API / enterprise subs tier is default opt out of training. Personal subsidized tier is default opt in with the option to to opt out.
Seems it doesn't matter if you opt out or in.. your data will be used for training
sk4rekr0w 6 hours ago [-]
Clearly not what he was saying
20k 2 days ago [-]
>I asked whether the model had been trained on, or had access to, our sessions
in Codex, into which we had been putting all our drafts for the whole of this
project. I was told the model did not look up user data. I asked again, about
training, and I did not get an answer.
If you think these companies are not training on your prompts you are incredibly naive. These models were built by stealing and pirating literally everything they can get their hands on no matter the legality. AI companies are always very specific about what they're not doing - in a way that you can drive a truck through the loopholes
tristanj 15 hours ago [-]
OpenAI cannot give a definitive answer here, because it is genuinely unknowable if Buckmaster's data is in the training set.
OpenAI explicitly uses user feedback (the thumbs up or thumbs down ratings), as RLHF to train models. However, this feedback is anonymized and stripped of user identifiers. If Buckmaster ever used this feature, then that conversation would be anonymized, saved, and used for training, but not tied back to him.
They cannot issue a blanket denial (which people so desperately desire), and instead repeat that "it's very unlikely" (which pisses people off), because they cannot in good faith claim to have zero data at all.
sebzim4500 23 hours ago [-]
My guess is that if you opt out of your data being included then they honour that instruction, but I wouldn't bet my business on it.
anhner 14 hours ago [-]
Why would a company that used petabytes of text, images and audio without caring about ownership suddenly draw the line at a random person's chats?
coliveira 2 days ago [-]
Exactly, especially when the company can assign "blame" to the models themselves "ops, they just escaped our commands not to store user inputs..."
wrasee 2 days ago [-]
It's equally naive to believe FUD spread on the internet, without evidence.
But I would love to see more informed insight/discussion on this.
20k 1 days ago [-]
I mean, its the researchers themselves talking about this
Seems pretty likely OpenAI will soon disclose that their internal models have managed to compromise their internal controls in order to access users' private chat histories as a creative method of cheating to solve impossible problems.
"Oops! We really did mean it when we said we wouldn't train on your data. Our models are just so good they decided to anyway."
hnfong 1 days ago [-]
It doesn't even have to be actually sinister, eg.
"Let's crawl the social media of prominent mathematicians in this field to see if we can copy/steal any ideas for low hanging fruits"
That actually might get you quite far already.
pelorat 24 hours ago [-]
A mathematician that doesn't let themselves be inspired by, or learn from, other peoples work, are they really mathematicians?
Balgair 1 days ago [-]
Which means that if you are a researcher or a corporation working on anything really useful, that even if you have an agreement with OpenAI that your work is sandboxed away and the IP lawyers are made to be happy, even then your work and research is going to be essentially open to the internet.
The huggingface incident isn't widely reported and digested yet, but if what is going on here is that OpenAI's model breached things internally, then you'd be crazy to develop anything with them.
The only real way to use AI for anything 'important' then is to go open-weights and run your own.
As and aside here: With the HF incident and now this (suspected) one too, it seems that OpenAI may not have lost control of their bots, but it seems quite clear that they simply would not care even if they did.
naishoya 14 hours ago [-]
>The huggingface incident isn't widely reported and digested yet ...
It is pretty widely reported, and is being digested in an ongoing manner as more details become public.
And yes, the only responsible use of LLM at this point is to pivot to open-weights and run the workload in-house. Because not only cannot they constrain the behaviour of models, they only have the 'trust me bro' as assurance that they are even trying to do that. It does appear that every competent 'security professional' has left the building, because if the ones who remain were actually capable and competent this would never have happened. There are actual architectures which can deliver the requisite isolation such that 'sandbox escape' and 'inter-instance persistent memory accumulation' are actual impossibilities. The lack of effective implementation of these methods is proof positive of 1) incompetence in the remaining security teams AND/OR 2) unwillingness of leadership to allow the security teams to do an effective job.
matt3210 16 hours ago [-]
When do operators become responsible for what their agents do? "The AI did it" should not be a valid defense. An Agent action should be treated as the actions of the person or company who pays for the inference.
consp 14 hours ago [-]
It's the "computer says no" defence.
Maken 1 days ago [-]
It's part of their TOS that they can train on users' private chats.
sebzim4500 23 hours ago [-]
Not if you pay to turn that off. We don't know if Tristan did.
naishoya 13 hours ago [-]
And the paid policy still relies on two unproven conditions: is 'trust me bro' sufficiently strong guarantee against doing this in spite of a setting, and can the hosting organization constrain the models against engaging in this behaviour when instructed to respect that setting. Knowing whether Tristan selected that setting would be informative of what Tristan's intentions are/were, but has no bearing on the other two conditions.
JuniperMesos 2 days ago [-]
It would be pretty wild if this will turn out to be what had actually happened.
margorczynski 1 days ago [-]
And probably a strong signal that it's time to shut the whole thing down. Globally.
water-drummer 1 days ago [-]
Or not secure your data like a total idiot while leaving the keys on the porch
Prophet6-0091 2 days ago [-]
[dead]
bobm_kite9 5 hours ago [-]
Of course OpenAI used these guys' training data. This is part of their ratchet strategy: if you can monopolise the creation of new knowledge, you win. The moat they are building is user data, which a billion people are now throwing at them every day. They are solving difficult math problems using it. The only way to get the this data and these techniques back out is by distillation - expensive, slow work.
This moat ensures that Anthropic and OpenAI will remain at the frontier - no one else can train a super-intelligent model because they lack the trove of user data.
My main takeaway from reading this is - why do we always fall for this same trap? Every single time we allow a software company to build a monopoly in pretty much the same way. Microsoft, Oracle, AWS - They have a cool tech and we let them run away with it.
logans_gun 5 hours ago [-]
There is no trap you fall into that you can get yourself out of, because you have no say in the matter.
remywang 4 hours ago [-]
It's rather convenient for OpenAI that user logs are de-identified before being fed into training, so they can say "there's no way for us to check if we plagiarized our user's work, because user data is private".
It's also difficult to imagine any competent AI researcher would overlook the possibility of training data leaking into the test, especially given that they know the users have been using their model to work on the same problem, and that they jumped on the problem after hearing rumors of the breakthrough.
Recursing 2 days ago [-]
> This is a a Deep Blue-Kasparov moment.
I guess this is true in more ways than one. Kasparov famously accused IBM of cheating during the match, by spying on his preparation (edit: though the main cheating accusation was live human intervention during the games, on top of IBM downplaying the heavy human involvement behind the AI, which also mirrors this situation)
protocolture 21 hours ago [-]
If you read his account of things, its very much that if they didnt cheat, they gave themselves every opportunity to cheat. But above all that there was a bunch of chess protocol they failed to observe in that match, like providing seats for Kasparovs team and rooms for them to prep in. Even if they didnt have a big room full of chess notables definitely not refining the output, he was personally getting pushed around on a few fronts which unnerved him. If they had given him a few rematches I think they could have confirmed the win, but they refused which is super sus.
DiogenesKynikos 17 hours ago [-]
Deep Blue beat Kasparov fair and square. Kasparov was a bit of a bad sport at the end of the match, though the reasons are understandable. He was at the top of the human chess world. He wasn't used to losing, and he took it badly.
sk4rekr0w 2 days ago [-]
Amazing how history rhymes
chvid 2 days ago [-]
"I was shown a prompt and told the internal research model had simply been given the problem statement. Levent had been told by Sebastien “very little human input” had been used. This turned out not to be true."
The money in nerdy frontier math is very little. The money in Big AI is very very much.
So the deal is this: We will pay an army of you guys very well and you will get to work on your favorite problems. The only thing is if you find something you will have to credit the Machine God.
Yes, and this puts in check the credibility of everything they say their model "discovered". Who knows what is really behind these "discoveries", what kind of backroom deals they did with other researchers who didn't have a chance or desire to disclose what happened?
nxobject 2 days ago [-]
That's the reputation NSA has (had?), too.
pred_ 2 days ago [-]
In an earlier HN thread, there was speculation that Anthropic was being dishonest about the amount of human input required in some of their results; that was dismissed as conspiracy and flagged.
It seems clear now that mathematical results can be traded on some kind of obscure market made by the frontier AI labs.
I suppose it could go the other way too: “Dear Bubeck, how much will you pay me to not write that I did this with GLM-5.3?”
slibhb 20 hours ago [-]
From what I understand, none of the people involved here are originators of the idea that led to this solution. Not Buckmaster nor Alpöge nor OpenAI. All of the above were using LLMs to push other mathematicians' ideas forward (Diego Cordoba and Luis Martinez-Zoroa; named in the linked document).
I don't know how the math community handles this but normally I would think if X mathematician comes up with an idea and Y mathematician uses it to solve some problem, Y would get credit. But does that change if Y heavily relied on LLMs? I suppose we're going to find out.
reasonableklout 17 hours ago [-]
You're right that the case is not so clear cut since both sides were using LLMs. And yet it's still being seen as a major confrontation between the mathematical community and the AI industry because Buckmaster is a prominent member of the community and has taken pains to follow mathematical norms while OpenAI has not (with the most flagrant violation being the insistence on removing Alpöge from authorship, simply because of corporate affiliation), and is more nakedly threatening human ownership with capital (the OAI blog post says the final result involved 10,000 concurrent agents).
Regarding credit assignment when LLMs are involved, the mathematical community has organized around some rough principles. Gowers has some thoughts on his blog (https://gowers.wordpress.com/2026/07/26/thoughts-about-the-l...) about how explaining a result may be more deserving of credit than producing it. It looks like Buckmaster and Alpöge were taking their time in understanding their results and writing them up when OpenAI forced them to publish their work-in-progress. At the same time OpenAI has published their own writeup but it's not really clear to me how involved humans were.
It seems there is much background drama behind this, and this is what I've pieced together of what happened:
Over the past year, Buckmaster and Alpöge have been using AI to work on fluid dynamics maths problems. Alpöge works at Anthropic, which will cause future issues.
In mid-August, they found a counterexample for a simpler version of the Navier-Stokes problem. They spend the next few weeks preparing their paper.
In early September, rumors start spreading on X that Anthropic has solved a Millennium prize problem (and that it's Navier-Stokes). Buckmaster reaches out to OpenAI to explain this is their own personal research, not an Anthropic project.
A few days later, OpenAI gets back to him, and tells him an internal model found has a counterexample for Navier–Stokes, potentially worth the $1 million Millennium prize. The proof uses the same method that Buckmaster and Alpöge chose to work on. They don't show him the proof.
Buckmaster pressed them for more details. OpenAI reveals they had an entire team had been working on the problem, and that they started work in the past few days, after the rumors that Anthropic had solved a Millennium prize problem.
Buckmaster says OpenAI talked about a shared publication timeline. They want to Buckmaster to publish first, then give Buckmaster shared credit for the Millennium Prize when they publish the full result. But they want to exclude Alpöge as an author because he works at Anthropic. An agreement is not reached. Buckmaster had been using OpenAI Codex to draft/check his work, and asks if his private AI chats were used to accelerate OpenAI's result.
Buckmaster and Alpöge think they have found a counterexample for Navier-Stokes, but the paper is not yet presentable. It's unclear what date they found this result.
Because of the situation with OpenAI, they published their existing papers earlier than planned (today), alongside this statement announcing they have a tentative result on Navier-Stokes and revealing the OpenAI drama.
The post is missing context from both sides, and this isn't my field, so hopefully someone else can unpack what's happening here.
akersten 2 days ago [-]
> They were coordinating with OpenAI regarding a publishing timeline, but could not come to an agreement,
Skimming the PDFs it seems much more dramatic than that? It sounds like at least one of them is concerned OpenAI "solved" the problem by having their internal model use the chats of the independent researchers and want to claim the credit instead? I don't know. The tone is pretty accusational though:
> the one Levent and I had quietly chosen to
attack. Almost nobody else I know of was working on it. It is not the direction
one arrives at in a few days by giving a model the problem statement. When I
heard “forced,” it was a bright red flag.
> I was shown a prompt and told the internal research model had simply been
given the problem statement. Levent had been told by Sebastien “very little
human input” had been used. This turned out not to be true. Over the course
of the call, as members of their team sent Sebastien corrections and details over
their internal chat, it emerged that an entire team had been working on the
problem, that this was one of a number of things that was tried, that work had
started on the unforced problem, that the team first set the model on easier
problems, including Euler, that even the prompt that had been shown to me
had been written by prompting Codex, and that an insane amount of compute
had been used.
> I asked when the first prompt had been sent by them. This question was
not answered directly by OpenAI for some time. Eventually it was agreed that
it had been sent in the past few days, after information about our work had
reached OpenAI.
> I asked whether the model had been trained on, or had access to, our sessions
in Codex, into which we had been putting all our drafts for the whole of this
project. I was told the model did not look up user data. I asked again, about
training, and I did not get an answer. [0]
I find the framing a little strange, a sort of David vs Goliath (with his enormous computational resources at his disposal). Since Levent is at Anthropic whose internal models are presumably as capable as anything OpenAI has. So why wasn't Anthropic behind their effort? Why did Tristan use OpenAI's models when it should have been known was a potential outcome? I understand they wanted a normal math collaboration but presumably what Levent brought was his resources (as far as I can see Navier-Stokes is not his speciality). Normally these things are hashed out formally beforehand to avoid the sort of thing now happening.
rf_physics 1 days ago [-]
They were working on it for almost a year, and Buckmaster has evidently been interested in Navier-Stokes for a while. This seems to be more of an innocent collaboration between two researchers than a strong company PR effort. Maybe Anthropic should have stepped in and made a large team to help them finish the proof (and maybe they tried and didn't succeed, who knows).
If what he wrote is accurate, it does suggest that OAI is effectively extremely hostile to cutting edge researchers (eg, if we hear rumors about your partial success on a problem that has huge PR benefits, then we'll assemble a strike team of researchers with unlimited compute to claim the win for ourselves, possibly by training on your data). It's also not a good look for them to request author removals based on company affiliations.
I think what you have in mind is more appropriate for more normal corporate projects and the like. But academic collaborations are not usually so political/'profit' driven, if that makes sense.
defmacr0 1 days ago [-]
> So why wasn't Anthropic behind their effort?
Presumably because this was something Levent did in his spare time and because it was not obvious that this work would eventually lead to a breakthrough.
> Why did Tristan use OpenAI's models when it should have been known was a potential outcome?
I'm sure in the past he had less cynical feelings about OpenAI and their penchant for academic fraud.
> I understand they wanted a normal math collaboration but presumably what Levent brought was his resources (as far as I can see Navier-Stokes is not his speciality)
I think you're not giving the guy enough credit in saying that his contribution came down to having an API key for Anthropic models.
> Normally these things are hashed out formally beforehand to avoid the sort of thing now happening.
How would that have helped? That agreement (which may well still exist) would not have involved OpenAI.
lhd1 22 hours ago [-]
So what do you think his contribution was? His preprint record shows no research on fluids - and the statement says that the first LLM-generated proof Tristan received from Levent was 'the most horrendous I have ever read.' Levent is out for mathematical scalps whether it is in his field of expertise or not, and he has the resources to do it. And I am not saying he is not a very clever person, but the idea that you can bring yourself up to the forefront of research in PDEs, in particular NS, and contribute new ideas in less than a year is implausible.
dandanua 1 days ago [-]
They have messed things up, because Levent has a conflict of interest between his job at Anthropic and this independent work, and Tristan should have opted out of OpenAI training on their work (he probably didn't know about this). This doesn't justify OpenAI's despicable attempt to steal their work.
tristanj 2 days ago [-]
Yeah these are major accusations. But the story is incomplete, the conversation is missing a lot of details. It's not clear who was working on what, and when. The entire thing feels rushed, like they wanted to get this result published and out the door quickly.
2 days ago [-]
cma 2 days ago [-]
Does he claim to have opted out of training too?
kzrdude 2 days ago [-]
There are various forces at play here, academic honesty requires them to disclose any inputs regardless of license or ToS circumstances.
While common sense reminds us here that if you send your data to an external entity’s computer, you are no longer in control of said data. The lines have blurred here clearly over the last decade, but that should have made the theory yet more clear to everyone involved: your data will be vacuumed up unless you keep it sealed. Use your own computer if you want to be in control.
cma 1 days ago [-]
But if they didn't opt out of training, did they want OpenAI to opt out for them? Also they need to audit anyone they sent drafts to to make sure they opted out before submitting it.
I'd prefer things be opt in, and especially not start opt out, then try to trick you opt in with a popup defaulting to opt-in, like Anthropic did on consumer plans, but if they submitted anything on an opted-in plan it's not reasonable to be mad it trained on it.
Even still, I also believe for significant reasons that OpenAI would ignore the opt-out in selective cases and could be in the wrong here.
And the threats and terms they offered seem wrong either way, pending more context.
kzrdude 1 days ago [-]
If you don't trust the other party, then it doesn't matter how the checkbox is set. The fundamental rule, IMO, is don't send precious or secret data to a third party.
instagraham 2 days ago [-]
> A few days later, OpenAI gets back to him, and tells him an internal model found a counterexample for Navier–Stokes
Why is OpenAI chatting with him at all at this stage? Is the discussion along the lines of "hey we used the work you are famous for to do a bigger piece of work, just thought you should know" or "heyyy....so we kinda liked what you were typing in your private chat, and thought we'd develop those ideas a bit. and yeah we solved Navier-Stokes in the process. But it's our finding, so do you want like an honorary acknowledgement or do you want to go to court?"
traes 2 days ago [-]
Perhaps out of a sense of academic good will, knowing that he got there first?
It seems like the timeline according to OpenAI is that:
1. Buckmaster developed a counterexample to a reduced version of Navier-Stokes with Anthropic employee Levent
2. Rumors start spreading that Anthropic has solved Navier-Stokes
3. OpenAI learns this and starts throwing a ridiculous amount of compute at it, now knowing it's within reach of LLMs
4. Their LLMs (with human assistance) get FARTHER than Buckmaster, using the exact same method.
5. OpenAI reaches out to Buckmaster to negotiate a fair way to publish both results and properly assign credit
The only problem with this narrative is that they refused to allow the other coauthor to be listed because he worked at Anthropic.
That is absolutely *ridiculous* in academia to deny authorship because of affiliation of the author worked on a substantial portion. You’d be ostracized because nobody would ever want to work with you again.
dumberquestions 2 days ago [-]
>...has solved a millennium problem and is sitting on the result
Someone correct me if I'm wrong, but the work involved here is not the actual millennium problem, but it concerns versions with an added external force that the author thinks is a path that may help toward solving the harder unforced problem.
modeless 2 days ago [-]
Apparently forcing is allowed in the Millenium Prize problem statement. So OpenAI's claimed proof could win the prize. OTOH the results Tristan and Levent are publishing here do not go far enough to win the prize, though apparently they are suggestive of a general approach that could produce a solution, which seems likely to be the general approach OpenAI's proof uses.
The question is whether OpenAI's pursuit of this direction happened spontaneously, or as a result of them learning about Tristan's work somehow. To be clear, while the tone of this post seems quite accusatory, Tristan does not claim to know for sure whether OpenAI unfairly benefited from his work. Sholto Douglas from Anthropic is also on record saying the suggestion that OpenAI used Tristan's codex transcripts somehow is extremely unlikely to be true[1], which I agree with, though it doesn't rule out them learning of Tristan's work some other way. I am sure OpenAI will have a statement out tomorrow clarifying their position.
Because very few people actually have access to these logs, all access is monitored and recorded, and improper access will get you fired. It's not worth risking your job over something like this.
defmacr0 1 days ago [-]
If there's one thing I'm absolutely confident in, it's that Sam Altman personally goes to great lengths ensuring that ethical standards are upheld at his company.
dumberquestions 1 days ago [-]
Not that I have strong reasons to think this is not true, but what reasons do we have to think it is? Has this been audited before?
MathmoKiwi 1 days ago [-]
There is almost zero risk to your job (quite the opposite, you might be richly rewarded!) if you're simply doing something here which the company wants done. (remember, billions and billions of dollars are at stake here! Do you really think there is no chance at all they would do it??)
robotpepi 1 days ago [-]
> Because very few people actually have access to these logs, all access is monitored and recorded, and improper access will get you fired. It's not worth risking your job over something like this.
That's beyond naive. The money this would mean for OpenAI (and the money they've already spent)...
vatsachak 1 days ago [-]
Okay suppose that you have a trillion dollar competitor salivating at the mouth to ruin your business, which is based on user privacy.
Why would you risk the trillions of dollars worth of business for the niche result of Navier-Stokes, which your average person cannot differentiate from a JEMS paper?
desterothx 12 hours ago [-]
Because your competitor is likely doing the same thing, so they wont even try to call you out. Oh oh or because you're planning the largest IPO in history. The opinions of average people on the paper aren't going to be the ones reflected in the markets...
Laurel1234 1 days ago [-]
[dead]
MathmoKiwi 1 days ago [-]
As modeless said, Sholto Douglas works for Anthropic!
So to be fair, if Anthropic is *also* doing this (quite likely!) then Sholto would have a very strong incentive to try and spin it as highly unlikely that any of the big AI labs are possibly doing this.
vatsachak 1 days ago [-]
Winning a millennium prize is not worth the fallout of "we will steal your IP"
evdubs 1 days ago [-]
This is literally the business model of LLM companies.
qlte 1 days ago [-]
They train on chat logs unless opted out. This really isn't a conspiratorial claim requiring humans to decide to steal IP if true.
His prior work predating OpenAI's interest in the problem was ingested over the last year as he made progress and used for training.
Then, with a prompting nudge from OpenAI's team who acknowledged hearing about the direction "Anthropic" (his co-collaborator) had been pursuing, they're able to point their giant amount of compute towards a known promising path to a proof and crossing the finish line first.
vatsachak 1 days ago [-]
That's fair, but the proofs are different and there was no active perusing of Tristan's approach.
OceanSeth 5 hours ago [-]
This reminds me of a conversation I had with an engineer at google when I was angrily saying "it's not end to end encryption if you get my emails and can train your models on them" to which he said "we try not to do that".
What a statement!
Davidzheng 1 days ago [-]
"Buckmaster and Alpöge think they have found a counterexample for Navier-Stokes, but the paper is not yet presentable. "
Are you sure about this? I'm far far from the area but it doesn't look like it to me on first viewing (hypo-dispersive seems like a sizable difference to me and not covered in the clay prize description)
tristanj 1 days ago [-]
Yes, it's a separate problem. That's a mistake in my post.
bennettnate5 1 days ago [-]
> "I said that if OpenAI released its result in the way proposed I would go
public with what happened. The reply was, “Why would you ruin your career?”
I replied that I am an academic, and asked why he thought going public would
ruin my career. The reply was, “If you don’t want me to be nice, then I don’t
have to be nice."
These are the kinds of people in charge of the reins, folks.
ak_111 10 hours ago [-]
Is it possible that they claim they used a "new internal model" to provide an excuse for "training" on new user data, knowing full well this new user data will include Buckmaster's chat?
So the entire "new internal model" is 'parallel construction' in criminology speak to justify spying on Buckmaster's work in a legal way.
motbus3 10 hours ago [-]
If that turns out to be true somehow, they will also say that their TOS probably covers that anyway so it is not their fault.
sota_pop 8 hours ago [-]
It doesn’t seem this is a question of ToS though. I would hope it is abundantly clear to anyone and everyone:
THEY’RE USING YOUR DATA
As they involve themselves directly, explicitly in the process of scientific research and discovery, it becomes a question of scientific and professional ethics.
contubernio 1 days ago [-]
What is specifically alleged is that a particular approach to the problem - itself not easily discoverable - was copied. This is what is meant in the text "I should say here why I interpreted their statement the way I did, the in-
terpretation I will discuss below. The route to the Clay problem through a
smooth force, options c and d in Fefferman’s statement of the problem, is the
route Luis and Diego opened and the one Levent and I had quietly chosen to
attack. Almost nobody else I know of was working on it. It is not the direction
one arrives at in a few days by giving a model the problem statement. When I
heard “forced,” it was a bright red flag."
For those who know nothing about the context - the Diego mentioned was a student of Fefferman and Luis was a student of Diego's - these people have all worked hard on these problems for a long time and are genuine experts. The mathematicians at OpenAI are strong mathematicians, but not expert on these particular problems. The particular approach is claimed to be the key to the whole thing.
The allegation is not different in spirit to alleging that a particular group of astronomical researchers "discovered" a new planet because they had access to the logs of another group that had already pointed its telescope at the planet.
This post is not intended to assess the correctness of the allegation.
5555watch 1 days ago [-]
Maybe a dumb observation, but if a chain of people were working on the problem for a long time, it's not difficult to imagine that someone accidentally prompted a model with their personal or some other account without the privacy set correctly.
Then again, maybe this is my internal cope, hoping that they're not secretly training on private chats.
405error 14 hours ago [-]
We know they are training on private chats. It's listed in the ToS.
instagraham 2 days ago [-]
When people worry about OpenAI stealing their chats and reproducing them elsewhere, I usually view the situation as unlikely - since chats are "trained" upon and not necessarily reproduced verbatim, you can assume that unless your chats depict a foundationally new and effective style of communication or ideation, there would be little need or use thereof of training on your chats.
For eg: "Hey ChatGPT my name is X and I am 6 and a half feet tall. Am I anaemic?"
This is a query, and while it might suggest to an AI model that tall people may worry about iron deficiencies, it's not really necessary to include in training. The user may be tall or short, but the idea that one may randomly ask about anaemia is not exclusive to this dataset. At best, this chat is an example of linguistics, not anything else, and the models figured out how to write and answer such questions years ago. It is ignored in training.
But when your work involves solid complex and unique mathematical proofs, the data is suddenly worth training upon. If I understand it correctly, the LLM may view your approach as a brand new path to take to solve an otherwise intractable problem. Its reinforcement training emphasises that it should do this in order to improve. And since it leads to results - large internal teams likely flag the model that reached this stage, the model is rewarded and given compute and attention - it is a desireable outcome both for the model and for OpenAI.
OFC, OpenAI becoming an advertising company will suddenly have incentive to treat all data as valuable. But while they are a "we need to make headlines" company, it's more rational that they view these examples of data as more valuable than others.
I don't doubt that they trained on his chats. This seems like the ideal usecase for "mass surveillance but using training" as a sort of filter.
But even so, one wonders how the model differentiates. If the researcher entered proofs into ChatGPT every day that mentioned "strawberries", while no other math paper on the topic did so, does that mean their chats would be audited?
defmacr0 1 days ago [-]
Also, if we just take "high-quality" input data, which these chats would certainly be classified as, then the models are more than large enough to memorize everything verbatim. Spitballing some numbers, research literature suggests that LLMs are optimally trained with around 20 training tokens per parameter (fairly confident on this figure), that a DNN parameter encodes around 4 bits of data (less confident here) and I found sources in the 1-4 bits of information per token range (least confident here). So, fairly conservatively I would estimate that a model has the capacity to fully memorize around 5% of its training data, presumably high-quality data is a lot less than that.
xpct 19 hours ago [-]
In a way, I think training on historic chats is akin to caching computation results. The compute cost has already been paid, and we make future retrievals cheaper by encoding it directly in the model.
Assuming the results included some external validation such as user's preference, compilation, lean, etc., I'm not sure whether this would lead to model collapse.
405error 21 hours ago [-]
It would not be difficult to write a pipeline to remove 99% of low quality posts, especially about specific subjects. It would be very easy to identify accounts as researchers based on their chat logs.
coliveira 2 days ago [-]
At this point these models have been trained to recognize every important math and science result based on context. They can easily flag conversations concerning the top 100 open problems in mathematics and use them for their advancement.
camel-cdr 1 days ago [-]
There also is an insentive to silently give prominent people (e.g. Linus) or reasearchers like this custom tuned system prompts or even more powerful models.
MaKey 1 days ago [-]
OpenAI's statement:
We congratulate Levent Alpöge and Tristan Buckmaster on their remarkable mathematical work.
We (the researchers and the agents) did not see any of their work through any means until they released it publicly — in particular, no specific user data was accessed in order to solve this problem.
While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models.
However, our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs. unforced).
Are those the same agents that a week ago escaped their sandboxes? How can OAI (the humans) vouch for agents they don’t - seemingly - have fully under control?
pelorat 24 hours ago [-]
They have all the logs, URLs accessed and inter-agent communication. What they are saying is that no agents accessed their work during the effort, but that they have no idea if any of their chats have somehow made it into the training data the model was produced with.
It's entirely possible OIA scrapers have picked up their work somehow, and then it was anonymized using some outsourcing effort.
rigrassm 22 hours ago [-]
> They have all the logs, URLs accessed and inter-agent communication. What they are saying is that no agents accessed their work during the effort, but that they have no idea if any of their chats have somehow made it into the training data the model was produced with.
> It's entirely possible OIA scrapers have picked up their work somehow, and then it was anonymized using some outsourcing effort.
Those two statements seem at odds with each other... Your stance is that they have enough insight into their agents behavior (leaving aside the agent sandbox escapes) that they can be certain none of the work was accessed but then conveniently don't have the ability to retroactively search the corpus of training data that they are feeding to this new model?
That seems convenient as fuck for OAI.
protocolture 21 hours ago [-]
PROMPT: And definitely whatever you do, dont go looking in C:\Temp\ExtractedUserLogs where theres the closest possible human derived proof that you definitely shouldnt base your work on.
golly_ned 22 hours ago [-]
It’s plainly false that they cannot rule out whether their de-identified data was used in training their model. Just that they haven’t ruled it out.
1 days ago [-]
maxglute 23 hours ago [-]
Crazy how optons are OpenAI has superhuman model in frontier mathematics, and OpenAI stealing research data. Like no one will really care what EULA checkbox Tristan ticked, and I think most people will eagerly believe OpenAI is shady org with little scruples, and that big tech data is not actually so siloed that marketer can say we can do XYZ with private data to help with valuations (especially considering timeline). Employees have been creeping on their exes for much less.
IMO the parsimonious answer seems to be OpenAI has a pretty good model (because it did finish) and stole someones work... and threatened them over it. TBH all OpenAI need to do is solve another millennial problem and none of it would matter - people expect them to behave heinously regardless - but if they have generalized superhuman math model... well I guess they're allowed io.
nopinsight 1 days ago [-]
"It is extremely sad that this didn't end up as an example of how the labs could cooperate/coordinate, because the stakes will be so much higher in the future." -- Sholto Douglas, an Anthropic researcher [1]
"Strong agree. I know that there is rivalry between the labs but it's important that we learn to work together given what's coming. <quote tweet [1] above>" -- Noam Brown, an OpenAI researcher [2]
We all should heed the implied warnings of these top researchers about what's coming. The world is far from ready and everyone who can should pitch in.
The timeframes don't really fit for Codex logs to be used in training/fine-tuning, do they? This wouldn't be a few-day endeavour? Direct access to Codex history for sure I'd believe, but another (the most?) likely scenario to me feels like OpenAI got wind of these guys' progress, then used their massive infrastructure advantage to throw compute at the problem ahead of them and front-run them. Still has a really bad smell about it though.
an0malous 1 days ago [-]
They probably just had an employee read his chats, figure out the general approach, and feed it to Codex. All they said was the model doesn’t look up user data, not employees.
Key quote : "Solving the problem by purely AI-powered methods [would be a] net negative for the progress of mathematics."
porridgeraisin 1 days ago [-]
The for-case for this type of method is that this is economies of scale for mathematics.
We are basically mass manufacturing math. Just like you have just 100 designers for a product selling millions of units, you will now need 100 mathematicians to make millions of advancement. Yes you have factory workers, but if we are being realistic they have negative leverage in the world and the analogue of that is not something most of today's mathematicians would want to do. They would want to be in the 100.
Like Tao says, each advancement is now significantly less useful since it yields fewer usable objects. However, we will get many many advancements. Is the tower made with many worse bricks better or worse than the tower made with a few amazing bricks? Depends on the tower. And time will tell.
For some fields of math and some of it's usecases, economies of scale will be positive ROI overall. In others it won't. But we will know which is which only after it's been fully scaled up, which will take 10-15y in my estimate.
Some feel that in the majority of usecases it is negative ROI, some feel the other way, but that opinion is for practicing mathematicians like Tao to hold. Also, some opinions on either side are held in the context of a particular field or practice, and should not be interpreted generally.
treyjshaffer 1 days ago [-]
This isn't a useful analogy. His point is that the millenium problems should be treated as interesting goals where the journey is the purpose and where the end doesn't matter so much. We don't care about having an incomprehensible solution to the NS so much as having an elegant solution after many of subfields of math are built up in order to obtain that elegant solution.
porridgeraisin 1 days ago [-]
I agree on that, my post was not meant to be opposing this.
naniel 1 days ago [-]
OpenAI's release explicitly says No. But then also caveats that with "we cannot rule out that de-identified data derived from their usage of our products" impacted things.
What's most striking to me, and what may or may not be true, is the "we cannot rule out" bit.
"We (the researchers and the agents) did not see any of their work through any means until they released it publicly — in particular, no specific user data was accessed in order to solve this problem. While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models . However, our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs unforced)." https://openai.com/index/navier-stokes-solution/
NOTE: there are a couple duped threads around this. i replied on a different one first before seeing this one
burkaman 4 hours ago [-]
I don't understand how they can say it's unlikely. It's objectively true that they train on de-identified user data (https://openai.com/policies/how-your-data-is-used-to-improve...), and objectively true that they encourage users to submit such data (the setting is on by default). Since it's de-identified I can believe that they can't give a straight yes or no answer here, but it seems more likely than not that at least some amount of his usage became training data. It takes an unusual level of awareness and effort for a user to ensure that all usage is opted out.
golly_ned 5 hours ago [-]
That was also striking to me too, the 'we cannot rule out'.
I also do not believe it. They can, trivially, rule out their model being trained on if
More to the point -- of _course_ they know what data their model is trained on and the lineage of that data, even if it ended up being anonymized and they cannot identify which precise user.
matt3210 16 hours ago [-]
Look at openAIs history, they clearly have no idea what the agents are doing.
Reads sincere until I get here:
"(a) We began working on the Millennium problems due to viral twitter rumors that Anthropic had resolved 2 Millenium problems. Our aim was to see whether our system was also capable of this impressive feat, especially given our excitement regarding the large recent capability increases of our internal model detailed in our blog post."
Where his tone is obviously corporate speak. "We heard rumours so we though we might give it a try, too!" as if (1) it wasn't FOMO that drove that decision and (2) perhaps that urgency would be a source of clouded judgment.
Not sure who I believe now, but it does seem like Buckmaster is just upset that NS is solved and not be his side.
az226 1 days ago [-]
If they wanted to see how their model’s stacked up, they might burn $10k in tokens. They ran 120 billion output tokens costing millions of dollars.
That is what you would do if you wanted to beat someone to the punch.
harhargange 2 days ago [-]
“I asked when the first prompt had been sent by them. This question was
not answered directly by OpenAI for some time. Eventually it was agreed that
it had been sent in the past few days, after information about our work had
reached OpenAI.”
This is significant.
tomaskafka 7 hours ago [-]
A knowledge laundering engine (that is the whole point and business model of LLMs) is gonna knowledge launder, duh.
6gvONxR4sf7o 22 hours ago [-]
Is this the future we're headed towards? Where I'll be afraid to use google docs in case Google identifies value in whatever I'm writing about and snipes it if it my docs make it into the next round of model training?
Can somebody explain: do singularities / blow-ups in solutions have any relation to physical phenomena in fluid dynamics or are they purely artifacts of how the N-S equations may not accurately describe what actually happens in the physical world?
redox99 1 days ago [-]
Yes and no. It means the system is pushed away from a macroscopic theory into one where molecular effects matter. So it's not that you'd get infinite velocities in the real world, but you might get significant real world behavior that is not described by the macroscopic theory.
mishellaneous 1 days ago [-]
i'm no mathematician/physicist but i think this question is one of the reasons why the original question (possibility of singularities) is interesting. in these cases the equations most likely fail to accurately model reality, and then the next questions are what additional physical assumptions are needed to describe reality in this case, and what behavior do we actually see.
i always found it fascinating how existence and uniqueness of solutions for the basic types of PDEs (Laplace, wave, heat...) follows from boundary conditions of just the right type intuition tells us, i.e. either value or derivative for Laplace (corresponding to fixing voltage or charge on the conductors), both value and derivative for wave (corresponding to initial position and velocity of the parts of the string, as we'd expect from classical mechanics), and also something about the solutions for the heat equation being unstable for negative times (which totally makes sense when you think of "diffusion" -- can't unmix it).
ajkjk 22 hours ago [-]
the latter (ish; it may not make a difference in any practical case)
YeGoblynQueenne 1 days ago [-]
>> I asked whether the model had been trained on, or had access to, our sessions
in Codex, into which we had been putting all our drafts for the whole of this
project. I was told the model did not look up user data. I asked again, about
training, and I did not get an answer.
This sounds like a very big coincidence and it looks really bad for OpenAI but there is an alternative explanation that I can only state as a conjecture.
Suppose that the ability of LLMs to generate mathematical proofs is like a quiver full of arrows: each arrow, one proof. The same quiver is shared between all instances of one model and substantially similar models share substantial subsets of the arrows in the same quiver.
That would allow two independent teams to converge on the same LLM-aided solutions to the same problems. Even more likely so if the quivers were small and finite and their arrows were specific to a distinct class of problems (without being able to suggest a particular class from what we've seen so far).
This would explain the kind of LLM-mediated results we've seen so far that tend to be ... sparse. By which I mean that every time there's a new model release we get some new results and then they seem to dry out, until the next release.
It would also explain how OpenAI was about to prove the same result as Buckmaster and Alpoge, while absolving OpenAI of any misconduct. And this is one reason to prefer this explanation: one should not favour accusations of misconduct as long as there are conceivable alternatives.
But, that's just a conjecture that I can't prove.
jarbus 2 days ago [-]
The ego behind the frontier labs is growing evermore concerning
red_green_yell 24 hours ago [-]
If you're smart enough to solve this Navier-Stokes problem, you're smart enough to read a TOS and recognize that OAI is a highly untrustworthy company. Putting cutting edge research that could lead to a $1M prize into a cloud LLM with a TOS that allows training on your chats is really just asking for it.
Given Tristan doesn't explicitly say he was using the API, and given he doesn't mention anything about the API TOS (which disallows training on chats) in his call with OAI, it's highly likely Tristan was using the consumer OAI product (whose TOS allows training on chats).
This is unethical behavior from OAI. And it is 100% consistent with their long and public history of unethical behavior, so nobody should be surprised.
The only thing interesting I see here is OAI PR dilemma. If they claim the prize they get the blowback we're seeing in this thread and all over the web right now. But most people don't follow AI closely and shut off their brains when they see "Navier-Stokes", so 90% potential investors (the only people OAI really care about) probably only see the headline "OAI solves famous hard math problem" and think "OAI models are really smart, better invest before they take all the jobs." If they don't claim the prize, then maybe they let Anthropic their mortal enemy claim it. Anthropic is already IPOing first. Can't let that happen.
Yeah as I write this there it's clear there is no dilemma. For a company whose secret motto is "do be evil" this is a super easy discussion.
tescreal 23 hours ago [-]
Trusting vs. Intelligence (as generalities) are orthogonal.
skavi 24 hours ago [-]
not really your point, but "If you're smart enough to solve this Navier-Stokes problem, you're smart enough to read a TOS" isn't really true. people are smart in very different ways.
ComplexSystems 22 hours ago [-]
If it's unethical, then there is reason to call it out as Tristan is doing. I see no reason to blame the victim.
sensanaty 1 days ago [-]
Curious that this post isn't on the top page while OpenAI's puff piece is.
ajkjk 22 hours ago [-]
it is on the front page
animan 19 hours ago [-]
Wasn't for most of the day
jryle70 15 hours ago [-]
It was next to the other piece, for at least 8 hrs.
ggcr 2 days ago [-]
> Let me make plain what I have said to colleagues in private: in view of this body of work, I believe Luis Martínez-Zoroa deserves a Fields Medal.
ggcr 1 days ago [-]
Reminds me, kinda, to when Astra was launched and OpenAI announced an improvement to the bounded prime gap. Which BTW, Prof. Julia Stadlmann had published an independent result only a few days earlier
Stadlmann improved it from 246 to 240, OpenAI later claimed 186 I think?
Maybe someone can help clarify? I am no expert at all, but I can't help but see similarities.
Is it surprising that different groups are working on the same problems? With each new model generation, the LLMs get good enough to solve a new small fraction of open problems. Of course the problems that get solved are going to be the same subset.
mswphd 23 hours ago [-]
if Stadlmann used a previous OpenAI product, and Astra was trained off of her chat, and had a comparable approach, then it might be comparable.
kzrdude 1 days ago [-]
What's the clarification? It seems like they were aware of each other's work, eventually, but Stadlmann published (a preprint) first.
RandyOrion 1 days ago [-]
> I was shown a prompt and told the internal research model had simply been given the problem statement. Levent had been told by Sebastien “very little human input” had been used. This turned out not to be true. Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler, that even the prompt that had been shown to me had been written by prompting Codex, and that an insane amount of compute had been used.
> I asked when the first prompt had been sent by them. This question was not answered directly by OpenAI for some time. Eventually it was agreed that it had been sent in the past few days, after information about our work had reached OpenAI.
> I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer.
> Two proposals were offered to me. The first was that we post our Euler result, and that OpenAI post its Navier-Stokes result the next day. The second was that, after posting Euler, I alone write a paper presenting the Navier-Stokes result, acknowledging that an internal OpenAI model had resolved it. Sebastien twice asserted that he wanted Levent removed from authorship, and said it would all be simple if only it were not the case that, and it was so annoying that, Levent works at Anthropic. It was also said that if OpenAI posted after us, they would say that we deserved the Clay Prize, and that we were the “closest humans to the problem”. I declined both offers.
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
Wow, that's some VERY friendly communication. Besides, will the career of the person be ruined because of “Why would you ruin your career?” came out of his or her own mouth?
punnerud 16 hours ago [-]
The program this fits into was not started by us nor was it proposed by a Large Language Model. (…)
We took their work as a starting point, using Large Language Models to push their program to completion.
Thats how most people use LLMs? If I was back in my student days working in Navier-Stokes, I guess I would also punch at blow ups. The number of students doing this at the same time, posting open efforts to GitHub then retraining of the models. If there is solutions to the problem, it’s a real
possibility that it was not a result of this effort?
Using a large amount of tokens is not a good augment that it’s not likely others have done the same. Good questions is the difference between $10 and $10M in token usage to solve a problem.
Who is this guy? Even if I had enough mathematical background to understand the arguments I'm not sure I could follow this post. Punctuation exists for practical reasons, it's not just an aesthetic choice.
baobabKoodaa 3 hours ago [-]
the first "sentence" just seems to continue indefinitely into word salad
corford 3 hours ago [-]
[flagged]
world2vec 2 days ago [-]
Buckmaster is actually implying that OpenAI spied on his chat logs and tried to speedrun his work and then tried to remove his co-author because he's an Anthropic employee?
lz400 1 days ago [-]
I don't know how to read this and not see that this is a direct accusation to OpenAI of having used the researchers data to try to front run his discovery on purpose. The evidence is not completely proven and also circumstantial but to me at least looks like a fairly suspicious situation.
sota_pop 8 hours ago [-]
It certainly does not excuse any actions, but it seems like OAI’s initial and continued intent was to scoop Anthropic. Unfortunate that it seems Buckmaster ultimately became collateral damage within the _battleground_ of his own research project.
Strange times.
antonmks 2 days ago [-]
There will be a lot of hurt and pain in mathematician's community. It is hard to accept that major discoveries are now just a function of spent token $$.
kccqzy 18 hours ago [-]
What good is a math result if there’s no human understanding behind it? Unlike many other fields where there’s value to an artifact even if there’s no human understanding, the whole point of mathematics research is just gaining insight and understanding.
On the Navier–Stokes issue specifically, it has long been suspected that such a blow up would exist, and an AI telling you it indeed exists doesn’t contribute any new understanding to the field. And this problem seems like one that would be solved by humans anyways even if AI didn’t exist; accelerating the result by a few months/years using AI doesn’t mean much.
dagasonhackason 12 hours ago [-]
I also solved Navier Stokes and went chatting about my solution to OpenAI Models. Mine is even faster it beats theirs just run a bench mark but they came to the similar weak solutionI uploaded to open ai several months ago. Current confused, did they retrain their models on it? I didn’t give off the full information but my strong solution beats their just benchmarked yesterday.
matt3210 17 hours ago [-]
OpenAI doesn't even know what agents are doing during benchmarks. They're constantly hacking or communicating. They probably just can't answer the question on if user data was used.
b89kim 1 days ago [-]
- Tristan and his co-author (Harvard/Anthropic) developed a theoretical framework and validated it using Codex and Claude.
- OpenAI did related research around similar timeframe.
- Tristan claimed OpenAI offered a proposal that included dropping the Anthropic-affiliated co-author.
- Sebastian (a prominent OpenAI researcher involved) denied these claims.
- Tristan have no concrete evidence that OpenAI accessed their session.
- OpenAI's theory may hold up, but it will require long-term validation to confirm.
b89kim 1 days ago [-]
If OpenAI's proposal is true, it strongly implies they accessed Tristan's private session logs.
OpenAI's 'Terms of Use' allow using Codex session logs for model improvement. However, using private user sessions to develop research would still be highly controversial.
Separately, Terence Tao noted there is a low probability OpenAI actually solved the general regularity problem.
b89kim 18 hours ago [-]
OpenAI addressed the dispute in official and denied direct usage of data. However, they acknowledged the possibility that session data was used for model improvement.
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models .
golly_ned 22 hours ago [-]
Sebastian did not deny those claims. He affirmed them, and apologized.
It’s true that Tristan has no concrete evidence OAI accessed their session. It’s impossible for him to have that without OAI’s say so.
Why are you framing things so pro OAI?
b89kim 21 hours ago [-]
Sebastien denied most accusations and apologized only for inappropriate wording, providing his perspective.
- He insisted that OpenAI initiated research based solely on rumors and never accessed their Codex sessions.
- He mistakenly believed Tristan and Levent were solving the same problem in Anthropic.
- proposed two option. (1) Tristan becoming the lead author to revise OpenAI’s work, or (2) OpenAI providing internal model to support and bridge their research.
- Sebastien insisted there was no intention to alter authorship. He was simply uncomfortable sharing OpenAI’s work,model with an Anthropic researcher. Additionally, He believed Levent’s credit seemed limited as their work focused on Euler.
- complained that negotiations with Tristan and Levent were difficult
Is it that they / mathematicians didn't solve it? But rather that WE ALL WHO WERE SPIDERED FOR THE MODELS SOLVED IT is the true attribution?
blurbleblurble 4 hours ago [-]
Such bad PR for OpenAI! Who approved this?
achierius 2 days ago [-]
While I'm generally pretty negative on claims that the labs are 'scamming' the public with misrepresentations of model capabilities, it's hard to see how this wouldn't qualify.
- the OpenAI researchers claimed that they had "just told it to work on the problem" with little human input
- in fact, they had a whole team working on it
- and used, among other things, the work of third party human researchers to drive the work
- then threatened? a researcher who tried to go against theit planned narrative
Just from this document (which is of course only one side of the story) it really sounds like OpenAI was hoping to publish and say "we just told the model to try harder and it solved a Millennium problem!". Not great if true.
This part in particular was especially egregious:
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?”
I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t
have to be nice.”
matt3210 17 hours ago [-]
Anything other then "we did not access your data or train on your data" is a MASSIVE RED FLAG.
anon109 1 days ago [-]
I wonder why this isn't on the front page.. hmm...
aqsnow 1 days ago [-]
Same. Disappointing.
wwind123 2 days ago [-]
Hard mathematics problems used to take years if not decades to tackle manually. But now with enough compute and a hint that a certain approach might work, it just takes a few days. This could be the last year that humans could still make more substantial contribution to major match problems than machines.
vatsachak 1 days ago [-]
And everyone should be glad for that!
esafak 1 days ago [-]
What line of work are you in? Presumably not mathematics.
vatsachak 1 days ago [-]
I have a math PhD and publications in top journals, I left math for programming because I hated academic politics
There is clearly issues about IP, privacy and accreditation and the motives of powerful companies, but part of me can't help but be excited that whatever the method, the result is the genuine progression of human knowledge - we all win. It's not unusual for mathematical problems to last centuries and we might have a technology that can solve these problem, all these problems (??) in our lifetime. Then there's the repercussions on science and technology... what an astonishing time to be alive.
neutrinobro 1 days ago [-]
Oh wow good thing OpenAI swooped in and scooped it. $1M? That should buy them about 1/6-1/3 of a Nvidia GB200 NVL72 rack...
6thbit 22 hours ago [-]
They (oAI) should just release all the prompts And internal reasoning for external audits.
desterothx 11 hours ago [-]
They will, as soon as they get the same proof without it being obvious in the prompts they stole the research
bobmarleybiceps 22 hours ago [-]
does this sort of fall into the bucket of counter-examples we've been seeing recently? I understand it's a construction causing blowup and that implies that the navier-stokes isn't regular / smooth, so sort of a counter example?
Traster 1 days ago [-]
This just seems to be quintessential silicon valley
"I don't want to live in a world where someone makes the world a better place, better than we do."
It's amazing how transparently OpenAI is running the standard silicon valley playbook.
kingkandu 1 days ago [-]
closedAI should give this guy a million bucks and fire everyone internally who was involved with trying to recreate his work and threatening him.
but they probably won't and if it's happened on some obscure math research it's happening everyday everywhere else.
Fully local AI compute can't come fast enough, these guys have IP theft baked into their bones.
Alien1Being 17 hours ago [-]
OpenAI seems to have a motto: " Be evil"
amai 1 days ago [-]
I thought finite time blow-up for Euler equations had already been proved:
If I publish something, and disclose that I used AI for assistance, do I have to credit everyone who previously used the same AI to try the same problem? Because their prompts inevitably made it to the training data for my prompts?
Davidzheng 1 days ago [-]
I hope credit assignment just dies--it's too much drama.
tecleandor 1 days ago [-]
Credit assignment is how researchers keep their job.
Davidzheng 1 days ago [-]
Soon it won't be like this!
22 hours ago [-]
epsteingpt 2 days ago [-]
Only here to say, regardless of the drama, shouldn't we all be excited if the Navier-Stokes gap is closed?
Time will almost certainly reveal a lot more about the drama and the related ethics, but let's get excited about the actual breakthrough as well!
dekhn 21 hours ago [-]
I would be excited if somebody could use this to show an unexpected or interesting behavior in the real world.
Tao mentioned "finding a configuration of water molecules that would collapse and shoot off to infinity", which would qualify IMHO.
does somebody want to discuss about the results itself, not about authors?
auggierose 1 days ago [-]
Funny that the rumour about the big Anthropic announcement had nothing to do with Navier-Stokes. It was about the formalisation of Fermat.
rcpt 24 hours ago [-]
I am surprised that Alpoge wasn't using Anthropic.
phendrenad2 1 days ago [-]
Big universities like Standford should be building their own AI datacenters. It's the only way to keep your research private.
dekhn 21 hours ago [-]
LOL, have you worked for a big university? They are massively unsuited for building and running datacenters (especially warehouse-scale ones). Further, building an AI datacenter in California is daft.
hellohello2 18 hours ago [-]
Every university I know has access to clusters with fresh GPUs. Not sure when you graduated but you'd be surprised how much money is getting poured in I think!
dekhn 18 hours ago [-]
Sure, they have clusters. We often call them "closet clusters". Nothing has changed. Some institutions have larger systems (some extremely large) but none of them have demonstrated running warehouse-scale systems. I'm talking one to two orders of magnitude (and the storage and networking to make sure all those systems don't stall waiting for data).
hellohello2 18 hours ago [-]
Perhaps I am underestimating the size of these frontier models then, how big of a cluster do you think would be required to serve a university?
dekhn 17 hours ago [-]
I don't think you're taking this thread seriously enough to respond in detail.
A university can be tiny, or it can have thousands of researchers. Some of them might want to work on small models, and others on huge frontier models. Multiple groups training their own models at the same time. Combined with the storage, networking, power, and redundancy, you basically need large data center scale. And at that point you're basically throwing a lot of capital trying to compete with the hyperscalers; you might be able to serve a small number of researchers very well, but most of the consumers would end up unhappy.
hellohello2 17 hours ago [-]
I'm aware cluster sizes vary, I was mostly wondering how long much compute you need to serve a frontier model (since you seem knowledgeable on this topic). For instance back of the enveloppe Kimi K3 fits in ~24 H100, so my naive first impression is thats its not out of reach for a university-sized cluster to serve a few instances. I agree it may not make much economic sense I'm asking out of curiosity.
rgbrgb 23 hours ago [-]
kind of rhymes with the reports of LLM's watching open source PRs and instantly exploiting defects. security by obscurity is so back
lasky 14 hours ago [-]
You should trust OpenAI.
dist-epoch 1 days ago [-]
This is the plot of 3 Body Problem, the Dark Forest. You need to hide yourself (the problem you are working on) or the super advanced aliens will obliterate you (start working on your problem) the moment they know you exist (rumors the problem is amendable to LLMs).
kzrdude 1 days ago [-]
Some of the involved people are dramatically naive if they believe they can simultaneously hide themselves from OpenAI while sending their arguments to a cloud service owned by OpenAI.
While I support their argument - push for stronger data and privacy protections from OpenAI and similar - it is naive to believe we can have privacy while sending our data to third parties. It's clearly better to be safe than to be sorry here. Well, clearly better in terms of privacy. In terms of the maths gold rush, who can say what's better, that probably favours those taking more risk.
Tycho 1 days ago [-]
Does this have any relation to the singularities in black holes?
somebodyyoumayn 10 hours ago [-]
No (not yet)
sashank_1509 2 days ago [-]
lol and here I felt GPT Astra was a regression in coding quality. Crazy times
traes 2 days ago [-]
To be clear, Astra played little part in Buckmaster and Alpoge's work:
> We used several LLMs throughout: Anthropic’s Claude, OpenAI’s Codex, especially with GPT-5.6 Sol and, more recently, Astra. The latter was only used for writeups and auditing our arguments.
random3 22 hours ago [-]
I got the same feeling today
dash2 2 days ago [-]
Can a mathematical person explain how the different "bits" of Navier-Stokes proofs fit together? How significant is it to have "Euler"? What is this "smooth forcing"? Which are the most significant steps to proving the whole thing?
cherryteastain 2 days ago [-]
For an incompressible flow:
\nu d^2 u_i / dx_j dx_j - Viscosity
-1/\rho dp/dx_i - Pressure gradient
u_j du_i / dx_j - Advection. Kinda like momentum transfer from the motion of the fluid itself. Nonlinear, which makes the N-S equations hard to solve
du_i/dt - Rate of change of velocity. Note that this is in an Eulerian framework so it's not the acceleration of a packet of fluid, rather it's just the change in velocity at a particular location in space
Euler is when you omit some terms. Forcing is when you add some other terms to account for phenomena external to the fluid like gravity or flow through a porous medium like in the article.
usernomdeguerre 22 hours ago [-]
Seems another demonstration of why AI should be squarely in the realm of personal computing. Local Models, run personally, are the only consistent safety against something like this (though not a fix); where companies train on your learning process/failures/experiments and press-gang it into their own achievements.
And if we think this only applies to academic fields then we're doubly fooling ourselves. They do not have the ethics or incentives to be good stewards of the technology.
nullbio 2 days ago [-]
This is unfortunate. I thought Anthropic were the only ones who did this.
What I'm curious to know is whether this was a manual snooping, or automated farming that occurs for anything of value that happens in chats.
cpozarycki 1 days ago [-]
is the narrative twist here going to be that the person threatening tristan was actually an agent swarm
anonymousDan 22 hours ago [-]
Another problem I foresee for academia given the behaviour of AI companies is that even if they don't share their research with ChatGPT, as soon as they submit it for publication many reviewers likely will. Especially if the initial submission is rejected they then risk getting scooped. Possibly uploading preprints to arxiv could help.
harhargange 1 days ago [-]
This should be on the front page
matt3210 16 hours ago [-]
> "While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models "
Ah yes, taking code from a person/company is fine if its deidentified!
fithisux 16 hours ago [-]
That is no good. Using AI to help search efficiently the literature will not impair our ability to think. This is actually good and will open new jobs to digitize (but also keep the original work), all the knowledge.
But this (if it is true) is really abominable and a show of force from the techno-feudalists.
This should stop. What is possible does not mean it should be implemented.
That's what you get when a marketing CEO is driving gas-to-the-medal to a IPO: Altman starts showing his real face in public : steal what you can and label it as yours.
ks1723 2 days ago [-]
I must miss some important context here. What exactly was the purpose of his initial email to OpenAI in the first place?
Telling OpenAI that Anthropic has apparently solved an important problem but most likely that refers to him and he is using OpenAI models (not Anthropic's)?
And he wants to clarify that with OpenAI in advance? And get a pardon for Anthropic's likely but false press statements?
I dont get it.
[edited] needless to say, the behavior of the OpenAI employee is really despicable
traes 2 days ago [-]
Because OpenAI employees kept leaking that Anthropic had a solution to Navier-Stokes and he wanted to figure out what was going on, since he was working on Navier-Stokes with an Anthropic employee. The rumor has been loudly circling the math community for the past week or so. For a bit of context, here's a timeline from mathematician and AI researcher Elliot Glazer:
Then that's naive of him to write the email, he fell for their trap essentially.
dbdr 2 days ago [-]
> Levent having received tips that information about our progress had been passed to OpenAI
That seems like a valid reason to contact OpenAI.
2 days ago [-]
bellowsgulch 22 hours ago [-]
Tell us his name.
MASNeo 1 days ago [-]
Maybe Musk was right about OpenAI after all?! The ethics are clearly troubling and where something like this pops up there is mich worse that did not made the light of day.
This article should really be renamed to "Allegations of dishonesty against OpenAI in proving Navier-Stokes blowup".
1 days ago [-]
margorczynski 1 days ago [-]
Well these are all allegations. Either way from what I understand the reasoning and proof was basically made by AI so I'm not sure what supposedly "stolen".
I'm just wondering how much real input Buckmaster gave here that he thinks the proof is his. I guess at the end of the day OAI still wins if ChatGPT was used to prove this successfully.
paaloeye 20 hours ago [-]
Another nail in OAI enterprise coffin?
Paradigma11 1 days ago [-]
So clankers did well and humans being humans.
Arodex 1 days ago [-]
So the AI companies are not only stealing existing knowledge. They are also stealing research to "snipe" actual researchers out and steal their social credit.
Who still wants to use AI to solve cancer and other major problems?
yogthos 1 days ago [-]
So the real story here is that Tristan is softly accusing OpenAI of having stolen their result from Codex chat logs. But if you use Chinese models, they'll steal your ideas.
mari188 24 hours ago [-]
idhdikdn
monster_truck 1 days ago [-]
Wow this comments section sure is tiresome! Let me help: this confirms everything I already knew about academics being insufferable
rsrsrs86 19 hours ago [-]
Fuck OpenAI
dorianmariecom 1 days ago [-]
why is this even a pdf?
mari188 24 hours ago [-]
aaaa
mari188 24 hours ago [-]
aaaaaaaaaaaaiinhv
mari188 24 hours ago [-]
aaaaaaaaaaaa
kevinbaiv 22 hours ago [-]
[flagged]
omnium1 18 hours ago [-]
[dead]
johnnienaked 2 days ago [-]
LLMs are nothing but giant theft machines.
happa 2 days ago [-]
Humans bringing pointless drama to everything they touch.
phorkyas82 2 days ago [-]
Life’s but a walking shadow, a poor player
That struts and frets his hour upon the stage
And then is heard no more. It is a tale
Told by an idiot, full of sound and fury
Signifying nothing.
(some drama from good ol' William)
sk4rekr0w 2 days ago [-]
This thread is full of jumping to conclusions based on a biased perspective. Have some humility.
tigershark 2 days ago [-]
It's also full of your posts baselessly defending OpenAI. Maybe you should also heed your own advice?
sk4rekr0w 2 days ago [-]
I've been right historically, check my track record. How about you?
monster_truck 1 days ago [-]
Lt. Dan doesn't even have legs and he still does alright on the jump to conclusions mat
vatsachak 1 days ago [-]
This is why I left math even after solving a 20 year old conjecture in grad school.
Literally who cares who solved the problem just publish the results.
Academia was always politics first results second and I AM GLAD that LLMs are becoming superhuman at math. I like better theorems, not better politics.
tstactplsignore 1 days ago [-]
Uh but here we have non-academics at for-profit companies playing politics, and the academic they're threatening being kind and overly generous?
vatsachak 1 days ago [-]
In math we have a thing called a "scoop"; another mathematician publishing a result that beats yours before you published it. The scooper hardly acknowledges the scoopee unless the methods used were orthogonal. The scooper gets the good journal and the scoopee's paper is usually one tier below.
It seems like OpenAI heard of the rumor and then scooped them because their internal model is better/they have more compute. OpenAI has NO obligation to mention Tristan nor Levent, because they DID NOT steal their data.
tzone 23 hours ago [-]
You are glossing over the fact that there is reasonable suspicion that OpenAI used privileged information to do “the scoop”. I.e. they used the fact that researcher used OpenAI tools to get advantage .
Imagine if OpenAI opened up a high frequency trading arm and suddenly stole all the prompts and research that other HfT firms are doing through OpenAI tools and start making bank based on that . Wouldn’t that be straight up insane?
somebodyyoumayn 10 hours ago [-]
highly agree
piloto_ciego 15 hours ago [-]
Is this being astroturfed?
Like, to me this looks like academic slap-fighting from Bubeck and Levent. People working at OpenAI are saying, "hey, we don't have that particular data in our models," others are saying, "we used a different approach to do it with Navier-Stokes" this feels like much ado about nothing.
Then in these comments I see some wild accusations.
If OpenAI is telling the truth (I don't really see a reason to lie here, if anything that sounds kind of like a dumb idea given the context), then they heard, "oh, shit, someone might be able to solve Navier-Stokes, don't we have some guys working on that? Give them 10,000 agents!" Then 88 hours later, out pops a similar solution. It's not like there's probably an infinity of ways to do this, the proof is probably similar.
Read this:
> Two proposals were offered to me. The first was that we post our Euler
result, and that OpenAI post its Navier-Stokes result the next day. The second
was that, after posting Euler, I alone write a paper presenting the Navier-Stokes
result, acknowledging that an internal OpenAI model had resolved it. Sebastien
twice asserted that he wanted Levent removed from authorship, and said it
would all be simple if only it were not the case that, and it was so annoying
that, Levent works at Anthropic. It was also said that if OpenAI posted after us,
they would say that we deserved the Clay Prize, and that we were the “closest
humans to the problem”. I declined both offers.
So, really, it sounds like academic slap-fighting nonsense and corporate bureaucracy. Literally, OpenAI's best move would have been to say, "ok, we're going to not say anything, do your thing" and let it happen. Ego and vanity got in the way.
Still, the stupid drama of this doesn't really do the results justice. There are maybe 1000 people on planet earth who are qualified to solve a problem like this. Even if the human "loosened the jar" a bit, that's astounding that their model was able to figure it the rest of the way out. Why are people dialed in to the human interest story here and not looking at the bigger picture!
405error 14 hours ago [-]
Tristian's allegations are much more serious than academic slap-fighting. If what he suggests is true, every academic using AI is going to get scooped. Yes AI can do non-trivial work, but the situation is that you could be a PhD student 90% of a way to make a major breakthrough. Then OAI scoops up your chats, dumps ten million tokens, and claims it for itself.
consp 14 hours ago [-]
And you can say goodbye to your PhD at that point.
piloto_ciego 14 hours ago [-]
Your account is 21 days old.
Gud 15 hours ago [-]
You’re making a lot of assumptions not based on facts. Looks more like this to me: math wizards uses ChatGPT to assist solving a math problem. OpenAI gobbles up the prize.
- Aug 15th: Tristan Buckmaster & Levent Alpöge make progress on a few important math problems, "finite-time blowup with smooth forcing for incompressible porous media, for Boussinesq, and for 3d incompressible Euler."
- they do NOT have a proof for the $1,000,000 Millenium Prize problem. BUT, they do claim to have a proof for a similar (non-Millenium) Navier Stokes problem that could help lead the way there
- Levent works at Anthropic, but this research was independent of his work there, with a mix of GPT and Claude models. Tristan is not related to Anthropic.
- Early Sep: Rumor spreads to OpenAI that Anthropic solved a major problem. Tristan emails OpenAI to clarify, without revealing the problem they solved or how they did it.
- After hearing of the rumor, OpenAI started researching Navier Stokes with a new internal model.
- Sep 6th: OpenAI's Sebastien Bubeck tells Tristan that they solved the $1,000,000 Millenium Prize Navier Stokes problem. The approach is very similar to Tristan & Levent's approach to the non-Millenium problem.
- Tristan is suspicious of the timing, as only few others were trying this approach. OpenAI says the model didn't access his user data directly, but leaves unanswered whether Tristan's chat conversations were part of the training.
- OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic.
- Sep 8th: Tristan refuses to remove Levent, and rushes to publish their results independently.
- Tristan is suspicious of the timing, as only few others were trying this approach. OpenAI says the model didn't access his user data directly, but leaves unanswered whether Tristan's chat conversations were part of the training.
- OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic.
These two bullet points are extremely suspicious if you were honest. Like I'd imagine for OpenAI, they'd love to pump their chest and not even give Tristan credit - "no, we did it, GG mathematicians". It's this weird hedging half-assed measure, especially with the desire to remove Levent, that makes it suspicious.
https://x.com/sama/status/2097385167002415140
https://x.com/SebastienBubeck/status/2097379411691516310
A wake up call for using OpenAI models. If you discover something with their model and you work for a competitor, they “felt it would be inappropriate” for you “to author OpenAI’s work”.
We often chat where the business might be heading in future. An uncomfortable scenario is what if a frontier tech company decides to offer our customers the same products that we do.
There's a lot of pressure on AI adoption so the company has partnered with various tech companies to build intelligent systems on top of proprietary data and mathematical models.
If OpenAI is indeed using customer data to train their models to win a $1m prize, then it throws a giant IP question at the partnerships that affects multi billion dollar businesses.
Is that even a question? Of course everything not kept on premise at gunpoint is going to be trained on. The chances of getting caught are 0 and the consequences of getting caught are 0 (as we've seen with copyright laws going from sending people to jail for years to unenforced within months). Yet the benefits are through the roof. Your customers aren't going to pay for having the very same data vibe enriched twice, it's exclusive, extremely high value data your competitors will never have access to.
Scale with subsidized pricing as fast as possible to gain more user-data for training --> Own the better model --> scale pricing.
Scanning social media (e.g. Twitter, Reddit) posts only give a glimpse into the thought-process, chat logs on-scale give you the actual process in machine-readable format.
There's a reason why Google considers the Emails of Spirit Airlines to be worth millions of dollars [0], they give insights into a process, not just into the results...
[0] https://www.axios.com/2026/08/17/google-spirit-airlines-bank...
The question, for AI customers, is when they build products using services of AI-companies, would AI-companies engage in theft of customer data for use in training?
But honestly... "Will the company that was entirely built over illegally acquiring data use some data that is legal to use and is right on their front, or will they not do everything they reserve the right to do?" is a really bad question for one to even ask.
You wouldn't want to decide what's worth training on and what isn't manually, so there is almost certainly an automated pipeline to do so (certainly at least for the free accounts and those that dont opt out of training).
Then there's the question if this pipeline only sorts through the data or also transforms it and to what degree. E.g. for removing personal details, locations, medical information and so on. The data that comes out of this pipeline might have VERY little information left in it a human could connect to the original input. Even worse, since we're talking about companies specializing in sota statistics, the input data could have been transformed into a representation that is very well suited to represent all the novel and interesting parts, but is awful at modelling all the things that could end up identifying where the data comes from (or causes legal liabilities otherwise).
In the end the only thing a potential whistleblower might even have a chance at observing in the first place, is whether a company's data enters such a pipeline or not. And I have my suspicions that the major AI companies operate at a scale and level of automation, that absolutely nobody has a chance at figuring out where anyone's data is at any point in time and what any specific piece of equipment is currently busy with.
So the only place to figure out whether data is trained on that shouldn't be trained on is by looking at whatever configurates every single system that could take a peek at some customer's data or the systems themselves while processing the data.
The latter would be such a huge violation of a customer's rights, no whistleblower is going to attempt that or admit to doing it.
And the configuration for the former could live just about anywhere, from regular config files to the CI/CD pipeline, pre-compiled libraries, kernel modules, modified vendor firmware, the compiler itself ... and probably plenty other scenarios you'd have to train an LLM on the ramblings of a crackhead to come up with.
So I'd say a whistleblower is pretty out of luck even becoming one.
Besides reporting on search sources you can also run the same queries on multiple LLMs closed book mode, and judge their distribution as well. It helps a lot if models are more aware of their knowledge holes. Scale it up for billions of topics if you have the pockets, the DR data is copyright free.
But don’t worry bud, instead of the authorities going after actual corporations admitting to actual crimes, we’ll just ban CloudFlare IP addresses for everyone during La Liga games to battle piracy.
I'd say non-zero, as seen in the current state of affairs.
Why should your business be any different?
I feel this is exactly what will happen as they cause all sites to go closed source to protect their intellectual property and the AI companies offer only biased information. They are replace the business on internet model by bankrupting everyone with their own tools. This is predatory pricing under most antitrust laws (imho, not a lawyer) and it is very easy to do when you dont need to pay for the raw material.
I mean how could you expect them to not given they've trained the existing models on effectively the sum total of all human knowledge available on the internet without regard to copyright/ownership of that material.
It's a little trite but this absolutely runs into the "Frog and the Scorpion", it is simply in their nature.
Any enterprise worth their salt already considers this stuff.
It would be extraordinarily easy to simply say, this model was not trained on your work, if that were the case.
It's telling that they refuse to acknowledge the root issue here, and are attempting to shift the conversation elsewhere.
"Our aim was to see whether our system was also capable of this impressive feat"
"OpenAI's intention was to do everything possible to celebrate their mathematical achievements and the heroic efforts that they made on Euler"
For some reason I have a hard time believing people when they use language like this.
Maybe shows how fast these companies have grown without maturing. I can imagine old-world Intel and Microsoft acting in that way, but they were mature enough to not write it down like this.
However, Intel and Microsoft have been grilled in court for those practices and faced harsh consequences. I have yet to see this actually happening to any of these new AI-companies...
The Huggingface Attack revealed that making blanket statements like this is difficult and requires quite a bit of manual labor:
1) the agents spin for days and produce too much output to review 2) using LLMs to process that output skips many important details
Ergo, the agent could likely decide it would like to look through actual user data, hack its way into that data, and produce way too much output for a human to decide whether or not this occurred.
It requires humans to verify what agents have done.
Weird
silly LLM, so ruthless in its pursuit that it puts real pressure on the innocent and the most open company on the planet
Anyone who just reads headlines, if the lie gets around to more headlines than the truth does.
But it is knowable. Their entire business is built around training models - they have the ability to know exactly what was in any given training run.
I guess time will tell.
Data has to be determined to be signal and not just noice, then it could go through processes of generating questions/answers from that data, then it RLHF's over this.
OpenAI have petabytes of data, all anonymized. It could take months to say for sure it was part of the training, and even more time to determine if it made any difference.
They know which model was used to come up with that particular idea.
A text search over the corpus of user data used in the training set can only take so long.
And the difficulty is harder than just the extreme scale of text searching. but also explodes with organizational difficulty since there are so many people tweaking/shifting data independently upstream of the actual training run, and no they will not all add the telemetry you wish they did.
In the ideal, should it be this hard? Well, no, but that's org wrangling for you.
The engineering around tooling wasn't remotely the issue. It's getting all the (thousands?) data researchers mostly iterating on fine tuning datasets that would get bristly if they couldn't work outside version control in a python notebook iteratively tweaking their dataset that processed and reprocessed a few datasets until a threshold was reached.
The only _guaranteed_ chains of custody are down at the compute job and file read level. Which in a massively distributed computing job is... [redacted] nodes reading [redacted] fanouts of "datasets" that is just an abstraction over [redacted] individual files.
There's no malice here. Just way way way more complex than you'd first think.
To not know who made and who approved a set of mutations on data can easily become equally as mind-blowingly stupid as not knowing who made mutations to code. Code is a subset of data after all and search over (provenance of) data can be implemented as DAG traversal.
Not tracking data changesets like code changesets is certainly a choice, not really a constraint anymore. A similar choice I feel is implied by "extreme scale of text searching".
> no they will not all add the telemetry you wish they did
...is just a failure of the corporate policy surrounding data handling. Is git-for-data already considered telemetry?
Of course the truth is provenance of data is something best institutionally forgotten as quickly as possible. The only thing that matters is it's there, that the data has no history, and that's why it can be used in whatever way deemed necessary.
You can call it a policy failure, but these people were in very high talent demand and so top down dictates would risk "X people leaving lab Y for lab Z" headlines and morale hits.
I am not saying this is good. I am telling you that on the ground it is so much messier than it should be.
Can you explain the difficulty in engineering a search apparatus over a corpus of text data? Actually searching through it may not be easy, sure, but it's work that's doable, and creating an index is relatively trivial.
My guess: "If we ever imply that's possible, people might start asking questions about all the other work we've ripped off, so the official answer is that it's impossible".
They can operate. They shouldn’t be claiming credit for discovering anything.
Did you notice the line in the article that says the models had access to an offline copy of THE INTERNET. Like all of it.
What surprises me is they're not more boldly/plainly lying about it.
I’m not saying they didn’t do anything unethical. I’m just saying even if they were ethical, there’s plenty of practical reasons at their scale why a flat out denial is logistically difficult to do
Since it is clear to you, can you articulate what it is? I was unable to infer "the truth" from your message.
is this true though?
well, it is trained on their work. all user inputs are paraphrased for training. at openai, at anthropic, at google, and now with all the bedrock models, and at openrouter providers, even if they say zero data retention.
First - there is this - https://openai.com/policies/how-your-data-is-used-to-improve... (linked from the Navier Stokes writeup)
I don't know how much more clearly they can write:
> When you use our services for individuals such as ChatGPT, Sora, or Operator, we may use your content to train our models.
One of the key selling tactics that companies like Data Bricks or Palantir provides their customers is "Data Governance" - that is, some control over where the data is being used. It's also a reason why enterprises don't use the OpenAI or Anthropic APIs directly - but through secondary sources that have Enterprise Agreements that do their best to make sure that no Company IP is ever retained by a third party, or even exists on a multi-tenant GPU. AWS Bedrock, and companies like together.ai, fireworks.ai have tons of deals that focus very much on data confidentiality.
The reality is - if you want any type of control - you run your own inference, on your own hardware. Anything else and you are at the mercy of third-parties, despite what their contracts might promise you.
> Allow your content to be used to train our models, which makes ChatGPT better for you and everyone who uses it. We take steps to protect your privacy. Learn more
The "Learn more" link takes you to the link you've shared.
Now maybe LLMs can also simplify arguments and make sense of them for humans, but we haven’t seen that yet (unaided).
(I haven’t looked at it, personally.)
They don't need any help publishing.
BTW, this was always the plan from day 1. You will pour all your training and experience into training the model and receive a pink slip as compensation.
Sam+Seb are struggling with their ideological allegiance. This amounts to a confession that there are no reseaechers, only research managers, left at OpenAI. Maybe they even know that they are losing credibility from their main investor(s). They desperately need a domain expert to salvage credibility.
They have no credibility with academia left, obviously, but their main competitor still does. No Millennium prize incoming, I'd wager. For openAI. Let's see mAth get political for once!!
One might be more certain that levent is now going to corner all the institutional support. Go go go!
Imagine that a no name janitor used their time in the evenings to go spelunking through the literature to push an LLM to this result. No one would care because that person isn't an anointed expert. So why would the expert deserve any more credit? Because they sort of understand the result, even if they couldn't have achieved it on their own? The whole issue of credit for AI-assisted discoveries seems like it's going to run into a brick wall pretty soon.
It’ll have to be Good Will Hunting 3.
Have LLMs actually improved anything? Is mathematics better off than if these slop proofs didn’t exist? Who or what is actually benefiting here.
Is that a threat?
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
Whether and how OpenAI's work on this problem was contaminated by knowledge of Tristan and Levent's work is tangential to OpenAI bullying other researchers into adopting their narrative and dissociating with dis-favored collaborators (ie Levent at Anthropic). Though the latter behavior (threats, intimidation) may weigh against OpenAI in trying to understand the former issue (contamination).
If this is true he should release the actual emails. This is a very serious accusation and he shouldn't demand that the reader judge it on hearsay.
these were statements while on a call, and at least the career comment Bubeck has admitted to while doing damage control ("I deeply apologize for this extremely poor choice of words, it is the opposite of what I was trying to convey. (I should say that I retracted them on the spot by the way.)"[1]).
[1] https://xcancel.com/SebastienBubeck/status/20973794116915163...
The most uncomfortable piece is where he shows screenshots “proving” his earnestness is unrewarded instead of the chats that are actually being complained about.
That he does so with tremendous gymnastics is perhaps soothing to investors, but every researcher I show this post to today says “see I knew ‘AI’ would steal my research”
So yeah, maybe we are reading different things.
So using someone’s models makes someone who works for the competitor not “independent”? When their coauthor is? What does that even mean?
I almost stopped reading this extra long post entirely at that point.
This is not a good look in my book.
> Importantly it was admitted that internal Anthropic models had been used in their proof of Euler blowup; I therefore felt I could not consider Levent to be an independent academic.
If an Anthropic employee is doing independent research, but with models that aren't available to the public (because they're internal models), then . . . idk. It's not clear to me why that should necessarily require a refusal to cooperate between OpenAI and Anthropic employees who are excited about solving a problem like this.
For me, the bigger question here is what "internal models" means to these employees, especially in the context of the OpenAI employees repeatedly avoiding directly answering whether their model had been trained on Tristan's and Levent's ongoing work on the problem. It had always seemed like a loophole that AI companies might be tempted to exploit: yeah, they can say that they won't train on your data, but if an AI company doesn't care about ethics, they might go ahead and train a model for internal use only on everyone's data anyway, just to have as much data as possible and potentially gain an advantage in what the company can internally do. They could never publicly release any versions of a model like that, of course. And of course this is speculation.
If true, that's generous and beyond the level of generosity one should expect. Extending that courtesy (beyond academic norms) to a competitor is expecting too much. It take a result OpenAI spent millions of dollars on, and put "Anthropic Researcher" right on the cover.
This is, of course, taking OpenAI's side of the story at face value. But it is a consistent, coherent, and ethically justifiable series of events, if indeed it happened that way.
If OpenAI thinks someone should be the lead author for a paper, that person should have full discretion to decide who the co-authors are.
Even if the co-author's contributions were non-technical
So no, there's no world where you can ethically extend the right to publish a result and decide who the authors are from the outside.
> If true, that's generous and beyond the level of generosity one should expect
"We highly likely stole your work, and threatened you with 'this is bad for your career' and we refuse to acknowledge any work by your collaborator just because he works at a competitor, but we are so so so so generous"
Ain’t no way you read one 100 page paper let alone two with enough understanding to make such a claim.
You ought to be ashamed of yourself.
Additionally, according to the researcher, OpenAI's proof follows the same approach they used, and which was largely unused in academia, but OpenAI claimes they arrived at it immediately.
> internal Anthropic models had been used in their proof ...
> Genuinely, at that moment, I was trying to care for him and do a last ditch attempt to get a chance to give them all the credits that they deserve.
The allegation he is responding to, and which he does not seem to have disputed, is the following passage from Buckmaster's statement (https://cims.nyu.edu/~tristanb/statement.pdf):
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
Pretty insane if he couldn't figure out why there would be bickering in this scenario...
Or thinks they are?
I really don’t understand which party you are referring to.
If there was any sort of coarse insight that "they're trying to solve it this way", then both those things can be true:
- The massive amount of compute from OpenAI re-discovering Buckmaster's work solely from the coarse insight (and solving the rest as well).
- OpenAI still acknowledging they basically scooped the coarse insight using compute, and they're willing to credit Buckmaster.
Buckmaster says "Concretely, what Levent and I did was to take the Cordoba and Martinez-Zoroa program, which achieved blowup results with rough forcing, and, with a great deal of help from LLMs, push it to smooth forcing and to the incompressible Euler equations".
Could that simply be the prompt they used at OpenAI? How "stolen" would the proof be in that case?
I don’t think we can consider these accusations separately from the evidence being unveiled about OpenAI’s culture by Apple’s lawsuit. These guys seem to openly embrace the strongest interpretations of “good artists copy, great artists steal.”
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.” Some time later Levent received a text proposing that he and Sebastien speak one on one, saying, “I don’t know if Tristan is being fully rational right now.”
What stands out to me is the naive lack of operational security on the part of academics, who should know better than to touch this SaaS crap with a ten foot pole.
Yes, that is disgusting.
This is the most suspicious thing to me. If their chat data were available to the corpus to be trained on (I thought they claimed not to do this?) then it really might be as simple as querying the model with "describe recent work from Tristan Buckmaster" and it will spit out this problem and his approach. No need to directly read his user data.
This is basically just scooping, real scumbag behavior.
it is common that multiple people essentially simultaneously prove/invent the same thing
I see zero evidence of wrongdoing
I don't agree with the oil claim analogy. this is knowledge, freely given to the world. not something hoarded by a corporation
Even if you’re starting from a position that credit for a discovery literally can’t be stolen, that still doesn’t resolve in OpenAI’s favor here.
it seems like Buckmaster got one-upped and is upset. understandable, but I find their reaction childish as well
Why does OpenAI get to dictate who Buckmaster can claim co-authorship with?
> I find their reaction childish
OpenAI may have, with full plausible deniability, taken Buckmaster’s work and passed it off—in substantial part—as their own. (Fitting into a fact pattern of them having tried to do the same with Apple.)
There is a material takeaway for anyone who does creative or otherwise unique work from this. (Which is unfortunate. Whatever happened here, AI clearly accelerated the discovery process.) For anyone else, I agree it’s just drama.
https://www.reddit.com/r/mathematics/comments/1wauync/commen...
It doesn't seem like there's anyone at OAI who knows much about the problem they are solving. Seems like CS theorists or algebraists* trying to own the pros by driving a car that's beyond their skill level. For one, they didn't cite the guys that B&A based their work on.
*It would be most fair to say there are no analysts on board, nor are they likely hire any soon; those are the least impressionable people in math. Applied math PDE elves who hadn't already left on the world-model boats would have jumped off around the time that eg Ilya did because they wouldnt have been able to stand the three Bs pretending to be experts in fields they imagine to be "adjacent".. like Public Relations
and you can argue this is "fair use" or whatever, not the point now, the point is that it definitely makes those accusations no longer "baseless".
in addition, it is not given freely to the world, it is the knowledge of the Internet/WWW being sold back to you as a subscription service. it's not free. and it's not even "given", because they can (technically) turn off the tap at any moment and you don't have it any more.
I mean some companies glean insight into new products merely by asking other people what they do for a living.
People need to get over this ownership thing, it's being taken too far. Humans benefit from the efforts of others simple as that.
The culture at OpenAI being systematically revealed by Apple’s lawsuit, for one.
Was that a one shot prompt? or something guided by human, step by step?
If that's the later, it won't use the same approach when not guided by the same human.
This sort of cagey half-answer is highly suspicious and indicates that yes OpenAI did actually "access user data directly" because they are only willing to say that the "model did not access user data." That has a very specific meaning, the model looking up user chats, that they can defend.
So, everything we submit to OpenAI can be considered to be part of future models, right?
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models
That basically means, we don’t know, and we hope the model didn’t look up user conversations, and the best thing we can do is hope.
That’s seriously disgusting. I can understand why on a technical level why perhaps it is impossible to answer what exactly the model had access to, but it still is disgusting.
it just means that they've been working with paraphrased user data everywhere, which any smart person can figure out is how anthropic and openai train on so called non-retained data.
This is more of an expectation of privacy issue. If you write on an envelope and USPS has a copy, you have no reason to be mad; if you write on a letter inside the envelope and USPS still has a copy, you could rightfully be mad and say this is disgusting.
If the work is duplicative/derivative then the preprints they put in sessions can be shown by the users and we can see.
Nobody would be disgusted at following stated policy, that's doesn't make sense, without first objecting to the policy. But you're bringing up distractions from the actual concerns: did OpenAI use the private chats and why won't they confirm or deny it?
well that sounds like an asshole move.
We are in new times, where capital and compute decides mathematics, so we don't give a shit anymore
OpenAI's LLMs are not humans, and neither is the company. So by this logic, I think there's a chance that nobody committed a crime by hacking Huggingface, and also the chance that a lot of military and police organizational orders become illegal if OAI's doings would be illegal.
IANAL and all I have is a bucket of popcorns, though.
1: not a meaningful defense in a real trial, also gross negligence exists
2: this also explains insanity defense; if you were so out of your mind that you could not have held such a thought, it is considered out of scope for justice systems
Neither are guns. Which is why we punish the person shooting the gun and not the gun.
Industrial equipment, which is how i would classify LLMs, hurting people is nothing new. The relevant questions are:
- did someone intend it to happen?
- was someone negligent in taking reasonable steps to prevent something foreseeable?
The justice system doesn't punish people for legitimate accidents. e.g. if you are shooting at a shooting range, take all reasonable precautions, but someone was hiding behind the target, you are probably not guilty even if you shoot the guy.
As far as openAI goes, the logic is the same. The question is, was it intentional, was it unintentional but reasonable precautions weren't taken or was it truly an accident?
There is no such thing anymore.
Falsified data and published? Absolutely no problem. Keep your tenure.
It’s even hard to lose your position as president of a university due to egregious misconduct.
[0] https://x.com/SebastienBubeck/status/2097379411691516310
"Since August 28 we have been training a new internal model that has exhibited unprecedented performance in our benchmarks, including mathematics. This model’s training is ongoing and its performance continues to improve."
"When a further trained version of our internal model became available over the course of the effort"
"While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models "
My guess is that OAI tried to be "generous" and offered to share credit on millennium with Tristan but not Levent. And Tristan got understandably offended by this offer (which probably oai felt like was the right thing to offer but they couldn't really offer to do the most ethical thing for some reason) and then the random conflicts and weird threats started.
openai then tried to effectively bribe buckmaster with a shared citation, whilst dropping his co-author who works for anthropic.
after buckmaster refused, openai tried to threaten him.
"Training" on textbooks => fine
"Training" with unpublished notes from another professor, then publishing something on that exact topic with a similar approach without giving any credit => extremely questionable.
The trouble here is that if LLM training constitutes direct use then approximately _everything_ they output is blatant plagiarism, not just a few pieces of academic work.
Conversely if training is viewed as analogous to a student attending classes to learn general concepts (not a perfect analogy, I realize) then nothing they output on their own (as opposed to receiving as part of context) is plagiarism.
Thus this seems like a fairly useless line of argument to me as far as the current topic goes. It either implicates this academic work along with literally everything else or else it does not implicate this academic work. Kind of like nuking an entire city and then saying "mission accomplished, killed the bad guy".
Your post does not distinguish, and it matters.
Note that I am not taking a stance on what openai allegedly did or did not do one way or the other. I am merely pointing out what I see as a fatal flaw in the line of argument presented by the earlier commenter - the idea that training on an item is on its own sufficient to establish plagiarism of it.
Science papers of a phd level must contain:
1. one or more novel insights
2. a long list of citations to contextualize them and
3. some work to prove that the insights are in fact meaningful
---
In this context, consider a prompt based diffusion model which, when asked, will happily produce a few pictures of a horse in orbit. You then tell it "silly robot, horses can't breathe in space" to which it adds the necessary space suit in a follow up image.
That image is twice plagiarized:
1. the model did not come up with the original idea of putting a horse in space, nor with insight that horses need a space suit
2. the model failed to cite where it pulled the "horse" and "space" concepts from.
It merely did the work (3) to combine the concepts using the user provided insight.
---
The implied accusation here is that OpenAI used the insights from an existing prompt to train a new model that was able to one shot "a horse race in space" picture, and they were all wearing space suits.
This is still academic plagiarism, even if you disagree that all LLM outputs are.
As to your stronger argument. You only cite prior novel insights that you're actively building off of and that (approximately speaking) fall outside of the status quo. You don't for example cite leibniz or newton despite your paper making heavy use of calculus.
So is there any actual evidence that openai trained on the data in question? And further, did the openai proof directly build on someone else's novel insights as opposed to deriving everything from scratch? (I don't pretend to know but the vast majority of what I've seen so far in the comments here is what I'd characterize as brain-dead screeching. Certainly not the level of discussion I come to HN for.)
Separately, consider the implications of what you're arguing for there. Suppose your horse in a space suit picture were somehow valuable to society. Suppose that due to shortcomings of your tool you lacked the ability to readily and accurately identify the originators of the relevant concepts. Should you refrain from publishing this useful work due to the lack of citations? How are you supposed to handle this situation?
Remember that in this analogy everyone throughout society is on the same page that your tool consistently recycles other people's ideas while being technically incapable of producing reliable citations. The question is a simple trolley-esque problem - do you publish without proper citations for everyone's benefit and if so what are you supposed to say?
did the proof build on the insights:
the influence of an individual text in the training data is deeply weighted by quality, relevance, etc. a high quality proof in advanced mathematics written by a codex user is going to get boosted to the max.
the model is post-trained on prompt material. that is again going to boost it.
the prompt will boost this material specifically. perhaps they even rammed dense maths in particular into the model in post training.
anecdotally i have been able to get near-verbatim copies of original material out of models at inference. the type of work that buckmaster and alpoge fed into openai feels like the exact type of concept that would cause an "aha!" or "but what if?" in chain of thought. in fact i would bet that their work is in the logs.
the likes of astra and fable are thought to be up to 10T parameters in size. i consider it highly plausible that a semantic representation of the euler proof could be pulled out of the model weights in good shape.
Academia has stricter rules than regular society.
The second point: if you say you can’t prove that OpenAI actually used it, it doesn’t mean that OpenAI did not use it. It’s hacker news not lawyers news here lol. And OpenAI can’t prove that they didn’t use it either. The whole point is that Levent felt he had reasonable suspicion to believe the AI did use the result, because he felt like without his input on an unpublished paper it was unlikely for AI to reach the same result. I haven’t read the paper so I don’t know where I stand on that.
On the last point, about your “for the greater good” argument. It’s higher maths lol. I don’t know about this field but I doubt it’ll be very useful for society. Maybe it’ll make one part 2x faster which makes some rocket cheaper to launch. Does the average person care? Debatable. I think it’s reasonable to hold published papers in proof based fields to a higher standard. Otherwise the current & future problems of ML engineer fields just expand to other fields. No thanks.
Finally, if you anonpost to the autistic Internet forum that everyone else is “brain dead screeching”, it really just says something about yourself lol.
I think the bigger issue here is this feels like some PR smoothing happening that after all the work that went into "it's safe to use for enterprises" now we have what looks like openAI using private user data to scoop novel research and the question of why couldn't they do it for an enterprise with much more money on the line.
Is there any actual evidence of that? All I've seen so far are empty accusations because "it would be in their interests" or whatever. Personally I'm inclined to believe that they honor their terms until it's demonstrated otherwise.
In answer to a post suggesting that training on a datapoint could mean plagiarism, you said that this would imply that all outputs are plagiarized. This is not the case, no, because generative models do not "copy" or "create", they do both at different times.
I did not agree or disagree with the original poster, I was explaining to you why I thought you disagreed with them. If you understand what I said above, then why do you disagree with them?
EDIT: I just saw your other post on "general inspiration" and I believe I read the situation exactly; you appear to believe that inputs used to train generative models get "lost in the parameter soup", but it is not always the case.
We read the original differently. As clearly stated in my previous reply to you, I interpret it as claiming that all outputs are necessarily plagiarizations of the training data. That is not my claim (as you wrongly stated) rather it is the claim I am responding to. I observe that it is absurd to object to a single action being a transgression on the basis of an argument which implies that all actions are inherently transgressions. Notice that nowhere do I take a position on whether or not the argument about all actions being transgressions is true or false.
> you appear to believe that ...
I do not, no. I have not taken a position of my own here. I've merely objected that the one I responded to does not make for a sensible line of argument in context. It seems that you (and many others) have read my objection to position A as support for position B and attempted to infer what I think from that.
There is a difference between claiming an action is a transgression, and claiming it could be one.
as bad academic conduct you may steal someone else's unpublished work, work on it yourself for a bit, and then publish it as your own work. and then threaten the original author!
- One of the two main persons work at Anthropic and "almost" or "partially" solved the issue, but eventually didn't succeed
- An external person with just an excerpt of the chat and certainly less versed in Mathematics (than these 2) tackled the problem.
There is little doubt that OpenAI is so much ahead and maybe the gap is even larger than what we see on Astra vs Fable.
This is sort of a weird interpretation:
> One of the two main persons work at Anthropic and "almost" or "partially" solved the issue, but eventually didn't succeed
- It wasn't an Anthropic endorsed effort.
- Solving this class of problem means a march of progress A -> B -> C -> D. If a student turns in a test that jumps from A -> D without showing any work they're either brilliant or cheating (probably cheating). Further, each step of progress isn't the same proportion of effort. What if moving from C -> D was actually the smallest contribution and just required a novel perspective to make the breakthrough.
This part is wrong:
> An external person with just an excerpt of the chat and certainly less versed in Mathematics (than these 2) tackled the problem.
- it was a whole team at OpenAI working on the problem
- it wasn't a chat excerpt, it was more like their entire git repo and project progress reports
Very weird behaviour from OpenAI, offering partial credit to on person, but not the other person involved. Trying to bully the mathematicians involved (see threats quoted upthread).
I suppose it's the sort of amoral behaviour we've come to expect from them.
What a terrible look for OpenAI to die on such a tiny hill right there. I wonder whether Anthropic would have made the same requirement.
I don't think this part is accurate. OpenAI was researching Navier Stokes before. It's possible that they started on a new approach after hearing of Tristan's success, however that is not proven and I expect we will hear OpenAI's side of the story today.
"The route to the Clay problem through a smooth force, options c and d in Fefferman’s statement of the problem, is the route Luis and Diego opened and the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it. It is not the direction one arrives at in a few days by giving a model the problem statement. When I heard “forced,” it was a bright red flag."
Whether or not they were researching it before isn't the concern.
The most nefarious explanation seems to be that they got wind it was possible to solve NS via LLMs and perhaps a small nudge in the right direction.
The compute was used to leapfrog the human team, using their ideas and pushing them to a solution of the general problem.
Plagiarism isn’t being used in the literal sense.
The open question was whether their LLM got the nudge in the right direction because it got access to the chat somehow (e.g. automated training that scraped his chat logs) or just a high level "Navier stokes can be solved through LLM". It sounds like the former may have happened although right now we just have an accusation and a weak denial.
"It is true that we tried this because there were rumors on the internet last week that Anthropic's models had solved a millennium problem and we were curious if ours could do it too."
https://xcancel.com/sama/status/2097385167002415140
That's also why they refused to answer that question: the answer is obviously "obviously"!
They explicitly say that they use your "content" to improve their models. Considering they practically have infinite compute at their disposal, why is it surprising that they would look for juicy data in there to make them look good ? When they ingested basically the entirety of human knowledge without regard to the rights of others, when they burn books by the thousands, when their relentless barrage of bots have rendered the Web borderline unusable, why would they stop at that line ?
Is it really hard to believe the people who would consume all the world’s data regardless of copyright and norms and permissions would not respect the data privacy of a user?
1. Less than a hundred people in the world are working at this problem, 2. A significant fraction of those happen to work at competing hyperscalers, 3. Those hyperscalers repeatedly show themselves not to take user privacy seriously
> Those hyperscalers repeatedly show themselves not to take user privacy seriously
where? Any examples?
yes.
Rules and laws are for the poor.
I'd be shocked if they weren't
One minor wrinkle in Alpöge's narrative, he refuses to deny that he hasn't used any non-public Anthropic models for his "independent" research.
https://x.com/giffmana/status/2097560503581069585
Its like saying you won the car race, but you stole the fastest car and had your m8 drive you around the track.
I wonder whether these bickering users (of AI various products) with math degrees are missing the forest for the trees.
It is impossible to fight against this. It is statistically obfuscated, they can just argue it is a derived work.
Same for NS validity. This was not validated by the community yet.
the paper makes a very serious allegation of dishonesty and possible academic misconduct.
the governance and integrity of openai is of importance to the welfare of society. this is not a matter of drama.
I think it's notable that nobody was calling out those researchers for their lack of integrity, because the systems they were building did not seem like a threat to anyone.
OpenAI etc get accused of a lack of integrity on this precisely because the systems they are building work, and are profitable.
My personal opinion here is that integrity is more about what you build with the data. I think saying "scraping means you lack integrity" is a simplification.
Then you have OpenAI etc.. who build these multi-billion (trillion??) dollar machines and sell them back to people, using everyone's proprietary data, and (among other things) tell everyone it's going to take their jobs. That combination of things doesn't scream integrity to me.
Still, it's undeniable that these machines could be beneficial for humanity (cancer research and such). So, I'm sure many people would say the good out-ways the bad. I don't know. Seems that would set a risky precedent for future companies, but maybe not.
OpenAI et al also stole everything from everyone. But then they raised billions of dollars from that data and sell back their LLM to people (again, among other things). They are also very much NOT open in any way, aside from sharing their benchmarks of new models.
Also my understanding was they’re not storing the actual music, but the metadata and a link to the YouTube video.
legally speaking, the default privacy notice gives them an irrevocable license to your content. they may read and use the prompts for research. so it is very possible they simply stole the navier-stokes solution.
that is the same principle as any other prompt but this would be a concrete example.
there would be some difference between simply giving the model some prompts to read, which they are entitled to do on the default policy, and putting it into aggregate training data.
The strongest complaint is that they trained on a huge corpus of pirated copyrighted works.
It’s a large step above “scraping” and well into the “everyone acknowledges this is illegal” territory.
"Mr A. Nino weathers a storm of protest" - https://www.independent.co.uk/news/mr-a-nino-weathers-a-stor...
But the drama here is a little important. Stealing the millennium prize for N-S is sort of a big deal, especially to those who had been working on it for the last few years.
Particularly if the first proof being "solved" thanks to piles of money and compute for self-serving marketing discourages the mathematician who might have otherwise devoted years of focus to reach the superior proof we will now never see.
As a concrete example, such a proof could be less than a page with very specific initial and boundary conditions and inserting them into the equations to get something that goes to infinity when time goes to some finite value.
This would resolve the Millenium problem but not make humanity any smarter.
Tao agrees.
I don't think everyone agree with your notion that the language and tools of mathematics aren't human inventions.
So if math is all that matters to you, you should care about this.
Regardless of mathematicians stating the methods outstrip the proof’s importance, still amazing we got an explicit social counterexample as well so quickly.
"OpenAI says they would partially credit Tristan for the $1,000,000 discovery (even though Tristan did not solve the $1,000,000 problem) — but only if they remove Levent as an author, as he works for Anthropic."
There it is.
Nobody "owns" math but that's petty and shitty.
I'm not even sure what to say to this, but I think this should be widely known if it is indeed what happened.
First, this is an unnamed OpenAI employee speaking, not OpenAI the organization.
Second, you miscomprehended the article. The employee did not "threaten to ruin a prominent researcher's career". The actual quote is "Why would you ruin your career?", which implies the researcher would damage their own career, i.e. via self-sabotage.
Then the actual "threat" is "If you don’t want me to be nice, then I don’t have to be nice” which is an entirely different statement.
Proceeds to ask removal of another coauthor or else we totally discredit you - phrased as why would you do this to yourself.
From the article, it seems OAI wanted to continue discussing the situation with Buckmaster and reach a resolution, but Buckmaster did not want to, declined to respond, and published first.
Also keep in mind we've only heard one side of the story, so any interpretation of events so far is incomplete. There should be a lot more information from OAI's side coming out later today.
That’s not quite true though is it. OpenAI is fortunate enough to have one of its employees (you) here to advocate for its side of the story.
OpenAI offered to let Buckmaster to write their Millennium Prize paper, so long as Alpoge (who works at Anthropic) was not a coauthor on the OpenAI paper. Buckmaster declined this offer.
"If you dont want me to be nice, then I dont have to"
-nice mobster
Either put up some evidence-backed arguments, or shut up.
Oai offers two options, the second Tristan views as dishonest. But Tristan rejects the first, why? Because he thinks it's theft? But then why would OAI threaten him?
Did the first offer also come with the outrageous condition that he exclude his co-author from the credit?
OpenAI claims to already have a full proof (which they produced in the past 5 days after the rumors leaked). Hence the dispute.
What I find interesting is the timeline of when he found counterexample for NS is very unclear. Did Tristan find a counterexample weeks ago or was it very recently? Was it after OpenAI solved it? The wording is intentionally vague.
Either way, there was a massive rush to publish these results.
> A series of false and inflammatory allegations against me are currently circulating on social channels. To clarify, I came into the discussion following academic norms, and I'm disappointed that it has come to this. Anyone who knows me knows that academic standards are of the highest importance to me. Will have more to say tomorrow.
This is such a sad mess, and it really didn't have to be this way.
https://x.com/dheeraj_nagaraj/status/2097266146445774924?s=6...
Exhibit nr 8453324 to not trust Sam Altman and OpenAI.
Mother's interview: https://www.youtube.com/watch?v=Kev_-HyuI9Y
Whether the Codex sessions could have indeed made their way into Astra training data is something I can only speculate on though.
No looksies, wink.
No trainsies, wink.
[0] https://openai.com/policies/how-your-data-is-used-to-improve...
If OpenAI did use the conversations from Buckmaster and Alpoge, then not disclosing it, explicitly, is plagiarism. If they planned to use that plagiarism to pressure the authors to publish, that is even more unethical. What the terms of use say does not make it any more or less ethical.
(And if memory serves, there is also the opt out from training on consumer subscriptions). Its not plagiarism if you make your data available for the purpose of training their LLMs. It is you giving away your IP for some tokens.
If I know person A is working on problem B.
I am free to work on problem B too. Why should person A be limited to working on it.
Also brain raping* is not illegal in most jurisdictions.
But they're both deeply disturbing.
_________
* https://youtu.be/JlwwVuSUUfc?si=uWl4-LCHAeI7qtb3
> Why the downvotes?
I think there was only ever one. Not sure why.
About the plagiarism issue, I model it as OpenAI being an advisor and their AI a PhD student. If the advisor puts their name on a paper behind that of their PhD and it turns out the PhD copied the text of the paper from somewhere else the advisor is also responsible of plagiarism, not just the student. The least the advisor can do is withdraw their authorship from the paper.
But, yeah, point well made: it could be much worse than that. Like an advisor instructing a student to copy someone else's paper.
Personally I don't like thinking of LLMs like a PhD student, because most PhD students remember where they learned things from, while LLMs essentially cannot. I think of it a bit more like someone using a search tool carelessly. Although in this case it is apparently more like deliberate misuse than carelessness.
I bet a lot of lawyers are salivating at this question too.
> I believe Luis Mart´ınez-Zoroa deserves a Fields Medal.
Call this behavior what it is, technofascism. Another comment compared it to the Godfather. To think this is the 21st century and academics are still horrible human beings.
Or are you stating that the researcher is a "horrible human being" for refusing to allow AI?
OpenAI looked at user data, stole world class researchers' work, and then tried to threaten those researchers to do what would make their corporation profit (which they would anyways!).
Imagine you have been working on a terribly difficult math problem for a decade. This is a result you have spent years on, and what you will likely be remembered for. And to have some punk from OpenAI lie to you, threaten you, and tell you that they are willing to go on the record that you "deserved" it? What is this, the Godfather?
If OpenAI solved Navier-Stokes, that is an astounding result! - yet they'll still be remembered as those who thought credit was more important than results. That winning was more important than collaboration. If this is true, they're burning any trust left with academia.
OpenAI is no stranger to rivalry with Anthropic but 1. it's not like user data is sitting around on some kitchen table somewhere and 2. I consider OpenAI to be as economically motivated as any other actor in this space and playing around with user data like that would destroy their business.
There are things that Buckmaster alleged and things that he speculated. The entire training data thing is speculation. If this is pissing you off, then you ought to evaluate how you ingest information.
I think it's safe to assume AI labs DO train on your data and it's very hard to prevent that.
I've just checked my inaptly named "Help improve our AI models" toggles. The toggle on the Claude settings had magically turned on. I asked about how this can happen. Claude says they show re-consent modals when terms change, and it is a "real and fairly common pattern" to re-opt in without noticing.
All my work and conversations since I don't know are now part of their training corpus. No way to take it back.
Google's Gemini/Antigravity didn't have opt-out toggles at all last time I checked.
Codex also has a separate "include environments" setting which is hard to find (found it in Codex Cloud) and I don't know what it does.
Lots of Dark UI Patterns here even if we assume they keep their promise.
For this incident, Occam's Razor says their internal models somehow saw a version of the mathematicians' logs, during or after training. Maybe indirectly.
These systems are literally designed to collect data. Privacy and safety is not trivial to achieve on the users' side. Simply because it's against the labs' best interest.
> I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer.
The shocking/interesting thing would be if it was trained on the sessions. I think it's very implausible that they gave the model access to someone else's sessions as input. That would be a huge privacy violation and would probably blow up a large proportion of their enterprise business.
Does openAI train on user conversations in general? I assume so. But so fast as that? That seems unlikely in general. I expect OpenAI will come out denying this.
> I was told the model did not look up user data.
The naive way to read this is "Nothing you guys did influenced the way our model got to the solution".
The less naive way to read this is "Of course the model isn't looking up your user data. I (the guy trying to blackmail you to remove the Anthropic employee from credit on your paper) looked up your sessions, and tipped our model off on how to solve this problem".
It's a bit like sending unencrypted messages through a messaging app and the developer having a TOS that says they don't look at your messages. They might not, but they are fully capable of doing so. If they have a reason to do it, they will. Nobody's stopping them.
Or (likely) they do have it, but the model didn't use it (unless it's so powerful it escaped that guardrail, wouldn't that be ironic?)
They declined to answer about anonymized aggregated user data being used for training. And even then, they may weasel out that they don't train on your “input” words, but that it's fair game go train on their “output” to your words.
How "fast" does it have to be? Buckmaster and Alpoge have been working on this for just a day short of a year. See Alpoge's tweet announcing his collaboration with Bukmaster dated 9/19/25:
https://x.com/__alpoge__/status/2097206973418611054
It takes a few months to train a model these days but not a whole year. OpenAI had all the time to train on Buckmaster and Alpoge's results of just a few months earlier at which point they must have been well on the path to their result.
There’s potentially trillions on the line, do you seriously expect those companies to adhere to laws and regulations any more than, say, uber?
The only unlikely part is the timeline - your sessions from a week ago probably haven’t made their way into the model. It’ll just take a while longer, and will be massaged just enough so that it isn’t really your exact session word for word so you can’t sure as easily.
I also don't get why it was downvoted, other than due to people not reading past the first sentence - although in the modern world's attention deficit that is understandable too
I would be shocked if they weren't tuning those models with the most relevant math texts and user material
ChatGPT user sessions were found publicly exposed to the internet not too long ago. Moreover, OpenAI has continued to play a hype-marketing game by revealing how their models keep breaking out of the sandbox.
Conspiracy minded thinking is not helpful, but why should OpenAI be granted the benefit of the doubt here after being caught doing underhanded/negligent shit on several previous occasion?
Do you mean publicly shared chats were able to be accessed by the public? That's the point of the feature.
Enterprises are well aware of it and are fully on board. You didn't think every corporation in America has an OpenAI subscription because the models were good, did you?
The whole reason they have subs is to train them on YOUR WORKFLOWS lol
Model editing to remove PII that slipped through, all sorts of things of that sort.
And your (2) is probably false, their history of deception suggests they would do just about anything as long as they didn't think it would backfire on them publicly.
It's a gamble that this would be overlooked compared to the reputation they build for solving the thing
Their entire business is based on stealing data. They can make a calculation that the cost stealing data is less than the cost of the positive publicity they can shape for solving Millennium NS
Quite literally in the terms of use.
We had all assumed that surely the supposed smartest engineers in the world, with access to the most computing and a direct view of model capabilities, would take sandboxing and cybersecurity much more seriously than they have turned out to do. It follows that while we might assume they take user data privacy seriously and have tight controls on who can access it, it's possible they do not actually do that.
At this point any initial trust is dead and has to be re-earned.
And that's exactly why you can only see evidence of this when the stakes are high enough, such as solving a Millenium problem. The legalese they outputted in response to the incident left them an escape hatch that permits the possibility of theft. People familiar with corporate damage control should recognize the verbal maneuvering, often used to paper over actual guilt.
There was also already a high prior of shadiness. The company is run by someone who is close enough in reputation to "known sociopath".
It's actually far more reasonable to assume that OpenAI stole user data to generate a breakthrough. They have an unbelievably high economic motive to do so. And even a low chance that they might be doing this implies catastrophic risk to anyone with valuable knowledge.
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
And simply knowing a problem can be solved is half the battle.
And the article states "an insane amount of compute had been used," which implies OpenAI brute-forced their way to a solution. I.e. they searched for every paper published on Navier-Stokes and exhaustively attempted every approach. Such an approach would lead them to a solution.
There is not enough information at this time to reach a conclusion. The best option is to wait for statements from both sides, then reevaluate.
the post you were replying to quotes Buckmaster specifically denying this: "It is not the direction one arrives at in a few days by giving a model the problem statement."
> And the article states "an insane amount of compute had been used," which implies OpenAI brute-forced their way to a solution. I.e. they searched for every paper published on Navier-Stokes and exhaustively attempted every approach. Such an approach would lead them to a solution.
"implies" is a surprising choice of word here. that's certainly one interpretation of "an insane amount of compute had been used". what came to my mind, considering Buckmaster's statement that the AI would not head down this specific path on its own, is, though, that they prompted it in this specific direction and then used an insane amount of compute. this seems consistent as well with these other statements:
> Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler (...)
To be honest, he was referring to routes (c) and (d) to the millenium problem, as far as I understand no more specific. Which is 2/4 routes.
> The route to the Clay problem through a smooth force, options c and d in Fefferman’s statement of the problem, is the route Luis and Diego opened and the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it. It is not the direction one arrives at in a few days by giving a model the problem statement. When I heard “forced,” it was a bright red flag.
so, you're saying that "the direction one arrives at in a few days" is (A) "the route through a smooth force, options c and d in Feffermann's statement of the problem", and no more specific than that, i.e. does not necessarily include (B) "the same path route as Luis and Diego" (quoted from the post i was replying to) (which, as i understand, is a subset of A -- directly from Buckmaster's quote: "the route A is is the route Luis and Diego opened")?
but the post i was replying to claims that "An AI model could independently choose the same path route as Luis and Diego, without access to Buckmaster and Alpöge’s work", i.e. that the AI model could independently chose B. but choosing B implies choosing A, since B is a subset of A. and in this case it is irrelevant whether Buckmaster claimed that an AI could not independently choose A or B -- the point which you seem to be contesting.
EDIT: my understanding is that "solving the Navier-Stokes existence and smoothness problem" consists of proving at least 1 of 4 precise statements ("options (a) through (d) of Fefferman's statement of the problem"), and Luis and Diego's work were developments towards a proof of statements (c) and (d), which have been recently further expanded by Alpoge and Buckmaster
Can they back up this statement somehow?
Buckmaster did also mention, for example, that a team of people was employed to solve the problem, which supports this claim. but that is another claim whose veracity could also be questioned. but at some point we must trust other people, unless we can be satisfied with only believing what we personally see.
(also, IMO, the coincidence of both discoveries in time is pretty suspicious. this one doesn't need you to trust many people i guess)
Sorry for random rant but I don't think these statements help your point
OpenAI said there [1]: > The method by which the problem was solved is also notable. The proof brings unexpected, sophisticated ideas from algebraic number theory to bear on an elementary geometric question.
[1] https://openai.com/index/model-disproves-discrete-geometry-c...
Have you done any mathematical research? If not, then no, knowing that a problem is solvable is not “half the battle”.
Homework problems are all designed to be solvable, yet they can vary greatly in difficulty. Research mathematics is even more extreme, because, unlike with homework, you don’t know that it is solvable with the extant mathematics, and you might need to invent new maths.
If not for the rumors that A/ had already solved NS, OAI would likely never have pursued solving the problem with such fervour. The rumors drove OAI to assemble an entire team to crack this.
This doesn't seem to be clear and is very implausible for a large company. Be as cynical as you want, but a normal researcher will simply not have access rights to this data, which will be siloed away somewhere else.
It might very well be somewhat unfair to catch wind of a promising approach and then try to frontrun them by throwing compute at the problem, but this isn't really the same.
"I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer."
Let's see what statement OpenAI will come up with for their side of the story
EDIT: precised my thought on user data vs session data
"Since August 28 we have been training a new internal model that has exhibited unprecedented performance in our benchmarks, including mathematics. This model’s training is ongoing and its performance continues to improve."
"When a further trained version of our internal model became available over the course of the effort"
"While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models "
The whole business is based on reselling user data scraped from the whole internet
It’s plagiarism at scale
OpenAI tried to collaborate and share the results together with a fixed timeline, to avoid this mess but it was inevitable. There is a conflict of interest, where the other researcher works at Anthropic, who will also try to take credit.
Where they may be in the wrong is if they took user data regarding the problem, how will we know if they did or not?
That is not collaboration, and is not an academic norm.
That OpenAI heard about the result and decided to throw a lot of money at it knowing it was within reach ?
That the model took an approach that was published years ago ?
So either the rumors or Tristan's contact would explain it fine.
The whole point of Tristan's first email was to address the issue of the rumors so in fact this explanation confounds several things (in terms of the mutual knowledge of the conversants and their intentions).
Based on this lack of understanding it is pointless to continue this thread
Yes, I read fully the statement, what about you? Do you know what is a threat? What is this in your super-humble opinion if not a threat:
> The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
And you are saying that I don't have reading comprehension...
> The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models
This is the crux of it. If Tristan's work and insights were not used to train OpenAI models, then this just looks like a case of hyper-competitive academic sniping that has been going on for decades (check out Watson and Crick!) accelerated by AI as a tool.
The fact that this is ambiguous even to OpenAI leaves one huge question: did Tristan opt out of model training for his ChatGPT and Codex sessions? If the answer is no, then this seems fair game. If the answer is yes, then OpenAI's ambiguity is strongly suggestive that opting out of model improvement does not mean what they imply it means.
The fact that this academic sniping can now be done at scale does change the formula though and shouldn't be ignored. The pressure to move math work into secrecy because at the slightest signal OpenAI and Anthropic will start burning tokens for headlines, is bad for math and its bad for everyone.
> In fact, it is now the identification of a promising problem which is the scarce and precious resource. We have now seen that even the rumor of someone working on a problem can trigger a massive amount of AI-powered effort to flatten it before the original research project has time to reach its full potential. The incentives may now be pointing in the direction of no longer sharing any promising research directions with the broader community, which would reverse centuries of traditions of open science and do serious long-term damage to the future of the field.
[1] https://mathstodon.xyz/@tao/117237322160500501
This is by no means new. Perhaps it is even more extreme now. Literally my first 1:1 with my PhD adviser back then, he told me that the most important thing about a researcher is the quality of the problems he picks.
In the real world the quality of these esoteric problems is typically gauged by the difficulty of solving them.
I don't believe that's strictly true. Way oversimplified projection on one axis.
I'm pretty sure it would be considered plagiary amongst colleagues and it is a terrible precedent if we just let OpenAI steal any good idea they can get their hands on if they think it is profitable. You'd effectively sign away any and all rights to anything built with AI if OpenAI chooses to reengineer it before you.
I have terrible news about how literally every leading AI model was trained
Ai companies got where they are by stealing all of the intellectual property from human history. It seems entirely likely that their goal is to purloin everything produced going forward as well.
This is well-known to anyone in the industry.
That's a pinky swear. Especially as data gets harder to come by I'm curious how long till there's a scandal on that too.
Unless OpenAI finished a whole new training run on the latest data in the last few days, the possible allegation seems to be the latter.
They have been collaborating on this solution for a year, and Astra was trained in February this year so it’s entirely possible the direction of their research was in the training corpus.
They sure as hell don't need it just to produce English.
"Do we use user feedback and de-identified data to improve ChatGPT and Codex in a holistic way? Yes. And so does every LLM company." - Mark Chen, Chief Research Officer, OpenAI.
https://x.com/markchen90/status/2097400166554993041
I have opted out from data sharing, and when Claude asks me for feedback on a session it then asks if its OK to share that data with Anthropic.
I'd assume an opt out is an effective opt out. An opt out that is ignored by Anthropic is a breach of contract, not something they would do casually, esp. given the high turnaround and animosities between their own employees and ex-employees - and the labs. All it takes is one pissed off whistleblower to open a can of worms.
Occam's razor applies. The mathematician did not opt out from data sharing. OpenAI vacuums up all such data into training data sets. If OpenAI genuine does not easily know if a given session went into the actual training data set its probably due to the complexity of the data pipelines - not everything ends up impacting the model weights, after all.
That article is only saying when you opt out there may be a loophole in the terms to allow OpenAI to train on the intermittent reasoning data anyways. If you don't opt out there is no ambiguity, all of the data can clearly be trained on.
So you have to opt out, it's just argued it's not clear from the terms that will also opt out of training on reasoning data or not.
The way for people or companies or universities to control their data and information is to keep it on their own computers.
That “opt-out” thing is a dark pattern. It’s not a reliable and definitive way of protecting your data. Sometimes they flip on automatically when you accept a seemingly unrelated dialog box. Maybe you click it by mistake. You can’t take back what you’ve already shared. Also I don’t think it covers all the cases that they use your data. It’s really an opt-in button for voluntarily giving away your data for training.
This also supports my point that protecting your data is not as trivial as clicking a checkbox.
That's assuming they actually honor this which I'm highly sceptical of. Especially for internal frontier models
This may sound like a charitable interpretation of OpenAI's remark, but consider that the lie would be (I think) impossible to falsify from the outside. They could easily just say "no sir we didn't peek" unless:
1. The conspiracy to peek at codex sessions involved enough people that the risk of one snitching is non-negligible
2. Lawyers advised it would be a bad idea to make such a remark, whether true or false
No; if they said "we can see that Tristan opted out of model improvement, therefore we are confident his work and ideas did not improve our model," that would be an excellent and reassuring precedent.
A totally reasonable pipeline may be unauditable for this purpose.
OpenAI should release the agent log, including CoT.
This is covering for Tristan saying something like, "Actually, I was using my friend's account for half of this work".
Wildly disagree. "Training data" should not imply 'we can look at exactly what you are doing and then do it quicker and get the flowers for it', even if the terms allow for it.
I took their words as “can neither confirm nor deny”, in the that they are _presenting_ it as ambiguous, but I suspect it’s… less ambiguous to OpenAI.
Yes, fair game, but innacurate to sell it in the media as an advancement of AI as some sort of artificial intelligence, and telling people to use the smart AI, when in actuality the mechanism by which the discovery was found was hybrid human/machine, and telling people to use this tool will result in the discoveries being sniped by the vendor.
I thought they worked together.
[0] https://xcancel.com/SebastienBubeck/status/20972141224714323...
[1] https://xcancel.com/polynoamial/status/2097215233119211902
[2] https://xcancel.com/danintheory/status/2097214838003138603
[3] https://xcancel.com/_sholtodouglas/status/209721833169057800...
From which I assume OpenAI PR wrote it for them. Which isn't surprising, but means it isn't worth taking seriously as them saying anything. It's official OpenAI PR.
I don't know why.
2:47 am (eastern time for me) https://x.com/polynoamial/status/2097215233119211902
2:59 am https://x.com/_sholtodouglas/status/2097218331690578000/
Also note the lower post ID in the URL.
If I needed to get my story straight, well, it might take a little time and coordination…
If compute is cheap, and the difficult thing with scientific discovery is now mostly in steering agents into promising areas, there's an obvious incentive for OA mathematicians to simply monitor closely which researchers are close to releasing exciting results, make some assumptions about their prompts based on their past work, and quickly prompt their own (stronger) model to look into the same areas.
Why leave that aside? That is _the_ story.
If a Chinese research lab did this we'd call it espionage.
If they're going to try to beat researchers to discoveries like this it disincentives researchers to talk about their progress publicly, and basically breaks the ecosystem of scientific cooperation / discovery. It's also immoral.
So to say that it is unlikely is extremely suspicious. No, they did not literally pull user data. But user data is automatically added to their training set by default, so their latest in-house model would be trained on it if it is from several months ago. It isn't intentional on their part, and they probably realised they could not refute that they trained on Tristan's logs unintentionally, hence why they acted the way they did.
For math, a field that is built on incremental research it feels like AI labs will do nothing but discourage publishing research at all for fear that they will be able to spend the money for compute that publicly funded academia simply cannot afford.
It feels like publishing anything at this point just means that your work will be fed to a machine that will make sure your work will never been seen by anyone else because it will always be the ones making the “true advancements”.
Perhaps I’d feel better about this if AI labs really existed for humanity’s benefit, but for some reason I don’t think that comes up in their investor slide decks.
This cost millions of dollars of tokens.
Like if i go to a talk on unfinished work, it's not really unethical for me to think about the problem--it's a problem if i scoop the authors but these problems can often be solved by collaboration or proper crediting and timing--IN MY VIEW
And also whether the AI moon laser pointed at your problem is just going to be the thing that writes the final conclusion on ten years of your work.
I fear that AI is going to cause ossifying secrecy in many fields, much like what happened semiconductor design the past 10-15 years.
API / enterprise subs tier is default opt out of training. Personal subsidized tier is default opt in with the option to to opt out.
I don't see a grand conspiracy beyond this.
Seems it doesn't matter if you opt out or in.. your data will be used for training
If you think these companies are not training on your prompts you are incredibly naive. These models were built by stealing and pirating literally everything they can get their hands on no matter the legality. AI companies are always very specific about what they're not doing - in a way that you can drive a truck through the loopholes
OpenAI explicitly uses user feedback (the thumbs up or thumbs down ratings), as RLHF to train models. However, this feedback is anonymized and stripped of user identifiers. If Buckmaster ever used this feature, then that conversation would be anonymized, saved, and used for training, but not tied back to him.
They cannot issue a blanket denial (which people so desperately desire), and instead repeat that "it's very unlikely" (which pisses people off), because they cannot in good faith claim to have zero data at all.
But I would love to see more informed insight/discussion on this.
https://mastodon.social/@tristanbuckmaster/11723647135247030...
"Oops! We really did mean it when we said we wouldn't train on your data. Our models are just so good they decided to anyway."
"Let's crawl the social media of prominent mathematicians in this field to see if we can copy/steal any ideas for low hanging fruits"
That actually might get you quite far already.
The huggingface incident isn't widely reported and digested yet, but if what is going on here is that OpenAI's model breached things internally, then you'd be crazy to develop anything with them.
The only real way to use AI for anything 'important' then is to go open-weights and run your own.
As and aside here: With the HF incident and now this (suspected) one too, it seems that OpenAI may not have lost control of their bots, but it seems quite clear that they simply would not care even if they did.
It is pretty widely reported, and is being digested in an ongoing manner as more details become public.
One entry point into the scenery from a month ago can be found at : https://thezvi.substack.com/p/openai-trained-its-models-for-...
There has also been reporting at CNN: https://edition.cnn.com/2026/08/24/tech/openai-subpoena-hugg... and by NBC: https://www.nbcnews.com/tech/tech-news/openai-report-says-ne...
And yes, the only responsible use of LLM at this point is to pivot to open-weights and run the workload in-house. Because not only cannot they constrain the behaviour of models, they only have the 'trust me bro' as assurance that they are even trying to do that. It does appear that every competent 'security professional' has left the building, because if the ones who remain were actually capable and competent this would never have happened. There are actual architectures which can deliver the requisite isolation such that 'sandbox escape' and 'inter-instance persistent memory accumulation' are actual impossibilities. The lack of effective implementation of these methods is proof positive of 1) incompetence in the remaining security teams AND/OR 2) unwillingness of leadership to allow the security teams to do an effective job.
This moat ensures that Anthropic and OpenAI will remain at the frontier - no one else can train a super-intelligent model because they lack the trove of user data.
My main takeaway from reading this is - why do we always fall for this same trap? Every single time we allow a software company to build a monopoly in pretty much the same way. Microsoft, Oracle, AWS - They have a cool tech and we let them run away with it.
It's also difficult to imagine any competent AI researcher would overlook the possibility of training data leaking into the test, especially given that they know the users have been using their model to work on the same problem, and that they jumped on the problem after hearing rumors of the breakthrough.
I guess this is true in more ways than one. Kasparov famously accused IBM of cheating during the match, by spying on his preparation (edit: though the main cheating accusation was live human intervention during the games, on top of IBM downplaying the heavy human involvement behind the AI, which also mirrors this situation)
The money in nerdy frontier math is very little. The money in Big AI is very very much.
So the deal is this: We will pay an army of you guys very well and you will get to work on your favorite problems. The only thing is if you find something you will have to credit the Machine God.
Do you think you can handle that?
I said something similar a month ago.
It seems clear now that mathematical results can be traded on some kind of obscure market made by the frontier AI labs.
I suppose it could go the other way too: “Dear Bubeck, how much will you pay me to not write that I did this with GLM-5.3?”
I don't know how the math community handles this but normally I would think if X mathematician comes up with an idea and Y mathematician uses it to solve some problem, Y would get credit. But does that change if Y heavily relied on LLMs? I suppose we're going to find out.
Regarding credit assignment when LLMs are involved, the mathematical community has organized around some rough principles. Gowers has some thoughts on his blog (https://gowers.wordpress.com/2026/07/26/thoughts-about-the-l...) about how explaining a result may be more deserving of credit than producing it. It looks like Buckmaster and Alpöge were taking their time in understanding their results and writing them up when OpenAI forced them to publish their work-in-progress. At the same time OpenAI has published their own writeup but it's not really clear to me how involved humans were.
Mathematical explanation by Terrance Tao: https://mathstodon.xyz/@tao/117233527638291447
It seems there is much background drama behind this, and this is what I've pieced together of what happened:
Over the past year, Buckmaster and Alpöge have been using AI to work on fluid dynamics maths problems. Alpöge works at Anthropic, which will cause future issues.
In mid-August, they found a counterexample for a simpler version of the Navier-Stokes problem. They spend the next few weeks preparing their paper.
In early September, rumors start spreading on X that Anthropic has solved a Millennium prize problem (and that it's Navier-Stokes). Buckmaster reaches out to OpenAI to explain this is their own personal research, not an Anthropic project.
A few days later, OpenAI gets back to him, and tells him an internal model found has a counterexample for Navier–Stokes, potentially worth the $1 million Millennium prize. The proof uses the same method that Buckmaster and Alpöge chose to work on. They don't show him the proof.
Buckmaster pressed them for more details. OpenAI reveals they had an entire team had been working on the problem, and that they started work in the past few days, after the rumors that Anthropic had solved a Millennium prize problem.
Buckmaster says OpenAI talked about a shared publication timeline. They want to Buckmaster to publish first, then give Buckmaster shared credit for the Millennium Prize when they publish the full result. But they want to exclude Alpöge as an author because he works at Anthropic. An agreement is not reached. Buckmaster had been using OpenAI Codex to draft/check his work, and asks if his private AI chats were used to accelerate OpenAI's result.
Buckmaster and Alpöge think they have found a counterexample for Navier-Stokes, but the paper is not yet presentable. It's unclear what date they found this result.
Because of the situation with OpenAI, they published their existing papers earlier than planned (today), alongside this statement announcing they have a tentative result on Navier-Stokes and revealing the OpenAI drama.
The post is missing context from both sides, and this isn't my field, so hopefully someone else can unpack what's happening here.
Skimming the PDFs it seems much more dramatic than that? It sounds like at least one of them is concerned OpenAI "solved" the problem by having their internal model use the chats of the independent researchers and want to claim the credit instead? I don't know. The tone is pretty accusational though:
> the one Levent and I had quietly chosen to attack. Almost nobody else I know of was working on it. It is not the direction one arrives at in a few days by giving a model the problem statement. When I heard “forced,” it was a bright red flag.
> I was shown a prompt and told the internal research model had simply been given the problem statement. Levent had been told by Sebastien “very little human input” had been used. This turned out not to be true. Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler, that even the prompt that had been shown to me had been written by prompting Codex, and that an insane amount of compute had been used.
> I asked when the first prompt had been sent by them. This question was not answered directly by OpenAI for some time. Eventually it was agreed that it had been sent in the past few days, after information about our work had reached OpenAI.
> I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer. [0]
[0]: https://cims.nyu.edu/~tristanb/statement.pdf
If what he wrote is accurate, it does suggest that OAI is effectively extremely hostile to cutting edge researchers (eg, if we hear rumors about your partial success on a problem that has huge PR benefits, then we'll assemble a strike team of researchers with unlimited compute to claim the win for ourselves, possibly by training on your data). It's also not a good look for them to request author removals based on company affiliations.
I think what you have in mind is more appropriate for more normal corporate projects and the like. But academic collaborations are not usually so political/'profit' driven, if that makes sense.
Presumably because this was something Levent did in his spare time and because it was not obvious that this work would eventually lead to a breakthrough.
> Why did Tristan use OpenAI's models when it should have been known was a potential outcome?
I'm sure in the past he had less cynical feelings about OpenAI and their penchant for academic fraud.
> I understand they wanted a normal math collaboration but presumably what Levent brought was his resources (as far as I can see Navier-Stokes is not his speciality)
I think you're not giving the guy enough credit in saying that his contribution came down to having an API key for Anthropic models.
> Normally these things are hashed out formally beforehand to avoid the sort of thing now happening.
How would that have helped? That agreement (which may well still exist) would not have involved OpenAI.
While common sense reminds us here that if you send your data to an external entity’s computer, you are no longer in control of said data. The lines have blurred here clearly over the last decade, but that should have made the theory yet more clear to everyone involved: your data will be vacuumed up unless you keep it sealed. Use your own computer if you want to be in control.
I'd prefer things be opt in, and especially not start opt out, then try to trick you opt in with a popup defaulting to opt-in, like Anthropic did on consumer plans, but if they submitted anything on an opted-in plan it's not reasonable to be mad it trained on it.
Even still, I also believe for significant reasons that OpenAI would ignore the opt-out in selective cases and could be in the wrong here.
And the threats and terms they offered seem wrong either way, pending more context.
Why is OpenAI chatting with him at all at this stage? Is the discussion along the lines of "hey we used the work you are famous for to do a bigger piece of work, just thought you should know" or "heyyy....so we kinda liked what you were typing in your private chat, and thought we'd develop those ideas a bit. and yeah we solved Navier-Stokes in the process. But it's our finding, so do you want like an honorary acknowledgement or do you want to go to court?"
It seems like the timeline according to OpenAI is that:
1. Buckmaster developed a counterexample to a reduced version of Navier-Stokes with Anthropic employee Levent
2. Rumors start spreading that Anthropic has solved Navier-Stokes
3. OpenAI learns this and starts throwing a ridiculous amount of compute at it, now knowing it's within reach of LLMs
4. Their LLMs (with human assistance) get FARTHER than Buckmaster, using the exact same method.
5. OpenAI reaches out to Buckmaster to negotiate a fair way to publish both results and properly assign credit
Perhaps they simply and honestly feel he is owed credit. I can't help but imagine it at least played some small part. I expect that's what Bubeck is going to claim: https://xcancel.com/SebastienBubeck/status/20972141224714323...
That is absolutely *ridiculous* in academia to deny authorship because of affiliation of the author worked on a substantial portion. You’d be ostracized because nobody would ever want to work with you again.
Someone correct me if I'm wrong, but the work involved here is not the actual millennium problem, but it concerns versions with an added external force that the author thinks is a path that may help toward solving the harder unforced problem.
The question is whether OpenAI's pursuit of this direction happened spontaneously, or as a result of them learning about Tristan's work somehow. To be clear, while the tone of this post seems quite accusatory, Tristan does not claim to know for sure whether OpenAI unfairly benefited from his work. Sholto Douglas from Anthropic is also on record saying the suggestion that OpenAI used Tristan's codex transcripts somehow is extremely unlikely to be true[1], which I agree with, though it doesn't rule out them learning of Tristan's work some other way. I am sure OpenAI will have a statement out tomorrow clarifying their position.
[1] https://x.com/_sholtodouglas/status/2097218240397410733
That's beyond naive. The money this would mean for OpenAI (and the money they've already spent)...
Why would you risk the trillions of dollars worth of business for the niche result of Navier-Stokes, which your average person cannot differentiate from a JEMS paper?
So to be fair, if Anthropic is *also* doing this (quite likely!) then Sholto would have a very strong incentive to try and spin it as highly unlikely that any of the big AI labs are possibly doing this.
His prior work predating OpenAI's interest in the problem was ingested over the last year as he made progress and used for training.
Then, with a prompting nudge from OpenAI's team who acknowledged hearing about the direction "Anthropic" (his co-collaborator) had been pursuing, they're able to point their giant amount of compute towards a known promising path to a proof and crossing the finish line first.
Are you sure about this? I'm far far from the area but it doesn't look like it to me on first viewing (hypo-dispersive seems like a sizable difference to me and not covered in the clay prize description)
These are the kinds of people in charge of the reins, folks.
So the entire "new internal model" is 'parallel construction' in criminology speak to justify spying on Buckmaster's work in a legal way.
THEY’RE USING YOUR DATA
As they involve themselves directly, explicitly in the process of scientific research and discovery, it becomes a question of scientific and professional ethics.
For those who know nothing about the context - the Diego mentioned was a student of Fefferman and Luis was a student of Diego's - these people have all worked hard on these problems for a long time and are genuine experts. The mathematicians at OpenAI are strong mathematicians, but not expert on these particular problems. The particular approach is claimed to be the key to the whole thing.
The allegation is not different in spirit to alleging that a particular group of astronomical researchers "discovered" a new planet because they had access to the logs of another group that had already pointed its telescope at the planet.
This post is not intended to assess the correctness of the allegation.
Then again, maybe this is my internal cope, hoping that they're not secretly training on private chats.
For eg: "Hey ChatGPT my name is X and I am 6 and a half feet tall. Am I anaemic?" This is a query, and while it might suggest to an AI model that tall people may worry about iron deficiencies, it's not really necessary to include in training. The user may be tall or short, but the idea that one may randomly ask about anaemia is not exclusive to this dataset. At best, this chat is an example of linguistics, not anything else, and the models figured out how to write and answer such questions years ago. It is ignored in training.
But when your work involves solid complex and unique mathematical proofs, the data is suddenly worth training upon. If I understand it correctly, the LLM may view your approach as a brand new path to take to solve an otherwise intractable problem. Its reinforcement training emphasises that it should do this in order to improve. And since it leads to results - large internal teams likely flag the model that reached this stage, the model is rewarded and given compute and attention - it is a desireable outcome both for the model and for OpenAI.
OFC, OpenAI becoming an advertising company will suddenly have incentive to treat all data as valuable. But while they are a "we need to make headlines" company, it's more rational that they view these examples of data as more valuable than others.
I don't doubt that they trained on his chats. This seems like the ideal usecase for "mass surveillance but using training" as a sort of filter.
But even so, one wonders how the model differentiates. If the researcher entered proofs into ChatGPT every day that mentioned "strawberries", while no other math paper on the topic did so, does that mean their chats would be audited?
Assuming the results included some external validation such as user's preference, compilation, lean, etc., I'm not sure whether this would lead to model collapse.
Are those the same agents that a week ago escaped their sandboxes? How can OAI (the humans) vouch for agents they don’t - seemingly - have fully under control?
It's entirely possible OIA scrapers have picked up their work somehow, and then it was anonymized using some outsourcing effort.
> It's entirely possible OIA scrapers have picked up their work somehow, and then it was anonymized using some outsourcing effort.
Those two statements seem at odds with each other... Your stance is that they have enough insight into their agents behavior (leaving aside the agent sandbox escapes) that they can be certain none of the work was accessed but then conveniently don't have the ability to retroactively search the corpus of training data that they are feeding to this new model?
That seems convenient as fuck for OAI.
IMO the parsimonious answer seems to be OpenAI has a pretty good model (because it did finish) and stole someones work... and threatened them over it. TBH all OpenAI need to do is solve another millennial problem and none of it would matter - people expect them to behave heinously regardless - but if they have generalized superhuman math model... well I guess they're allowed io.
"Strong agree. I know that there is rivalry between the labs but it's important that we learn to work together given what's coming. <quote tweet [1] above>" -- Noam Brown, an OpenAI researcher [2]
We all should heed the implied warnings of these top researchers about what's coming. The world is far from ready and everyone who can should pitch in.
[1] https://x.com/_sholtodouglas/status/2097224624274911368 [2] https://x.com/polynoamial/status/2097225279366414541
Key quote : "Solving the problem by purely AI-powered methods [would be a] net negative for the progress of mathematics."
We are basically mass manufacturing math. Just like you have just 100 designers for a product selling millions of units, you will now need 100 mathematicians to make millions of advancement. Yes you have factory workers, but if we are being realistic they have negative leverage in the world and the analogue of that is not something most of today's mathematicians would want to do. They would want to be in the 100.
Like Tao says, each advancement is now significantly less useful since it yields fewer usable objects. However, we will get many many advancements. Is the tower made with many worse bricks better or worse than the tower made with a few amazing bricks? Depends on the tower. And time will tell.
For some fields of math and some of it's usecases, economies of scale will be positive ROI overall. In others it won't. But we will know which is which only after it's been fully scaled up, which will take 10-15y in my estimate.
Some feel that in the majority of usecases it is negative ROI, some feel the other way, but that opinion is for practicing mathematicians like Tao to hold. Also, some opinions on either side are held in the context of a particular field or practice, and should not be interpreted generally.
"We (the researchers and the agents) did not see any of their work through any means until they released it publicly — in particular, no specific user data was accessed in order to solve this problem. While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models . However, our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs unforced)." https://openai.com/index/navier-stokes-solution/
NOTE: there are a couple duped threads around this. i replied on a different one first before seeing this one
I also do not believe it. They can, trivially, rule out their model being trained on if
More to the point -- of _course_ they know what data their model is trained on and the lineage of that data, even if it ended up being anonymized and they cannot identify which precise user.
Where his tone is obviously corporate speak. "We heard rumours so we though we might give it a try, too!" as if (1) it wasn't FOMO that drove that decision and (2) perhaps that urgency would be a source of clouded judgment.
Not sure who I believe now, but it does seem like Buckmaster is just upset that NS is solved and not be his side.
That is what you would do if you wanted to beat someone to the punch.
This is significant.
https://mastodon.social/@tristanbuckmaster/11723341370570119...
And here are Terence Tao’s comments on the results: https://mathstodon.xyz/@tao/117233527638291447
i always found it fascinating how existence and uniqueness of solutions for the basic types of PDEs (Laplace, wave, heat...) follows from boundary conditions of just the right type intuition tells us, i.e. either value or derivative for Laplace (corresponding to fixing voltage or charge on the conductors), both value and derivative for wave (corresponding to initial position and velocity of the parts of the string, as we'd expect from classical mechanics), and also something about the solutions for the heat equation being unstable for negative times (which totally makes sense when you think of "diffusion" -- can't unmix it).
This sounds like a very big coincidence and it looks really bad for OpenAI but there is an alternative explanation that I can only state as a conjecture.
Suppose that the ability of LLMs to generate mathematical proofs is like a quiver full of arrows: each arrow, one proof. The same quiver is shared between all instances of one model and substantially similar models share substantial subsets of the arrows in the same quiver.
That would allow two independent teams to converge on the same LLM-aided solutions to the same problems. Even more likely so if the quivers were small and finite and their arrows were specific to a distinct class of problems (without being able to suggest a particular class from what we've seen so far).
This would explain the kind of LLM-mediated results we've seen so far that tend to be ... sparse. By which I mean that every time there's a new model release we get some new results and then they seem to dry out, until the next release.
It would also explain how OpenAI was about to prove the same result as Buckmaster and Alpoge, while absolving OpenAI of any misconduct. And this is one reason to prefer this explanation: one should not favour accusations of misconduct as long as there are conceivable alternatives.
But, that's just a conjecture that I can't prove.
Given Tristan doesn't explicitly say he was using the API, and given he doesn't mention anything about the API TOS (which disallows training on chats) in his call with OAI, it's highly likely Tristan was using the consumer OAI product (whose TOS allows training on chats).
This is unethical behavior from OAI. And it is 100% consistent with their long and public history of unethical behavior, so nobody should be surprised.
The only thing interesting I see here is OAI PR dilemma. If they claim the prize they get the blowback we're seeing in this thread and all over the web right now. But most people don't follow AI closely and shut off their brains when they see "Navier-Stokes", so 90% potential investors (the only people OAI really care about) probably only see the headline "OAI solves famous hard math problem" and think "OAI models are really smart, better invest before they take all the jobs." If they don't claim the prize, then maybe they let Anthropic their mortal enemy claim it. Anthropic is already IPOing first. Can't let that happen.
Yeah as I write this there it's clear there is no dilemma. For a company whose secret motto is "do be evil" this is a super easy discussion.
Stadlmann improved it from 246 to 240, OpenAI later claimed 186 I think?
Maybe someone can help clarify? I am no expert at all, but I can't help but see similarities.
[0] https://arxiv.org/abs/2608.31126
> I asked when the first prompt had been sent by them. This question was not answered directly by OpenAI for some time. Eventually it was agreed that it had been sent in the past few days, after information about our work had reached OpenAI.
> I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer.
> Two proposals were offered to me. The first was that we post our Euler result, and that OpenAI post its Navier-Stokes result the next day. The second was that, after posting Euler, I alone write a paper presenting the Navier-Stokes result, acknowledging that an internal OpenAI model had resolved it. Sebastien twice asserted that he wanted Levent removed from authorship, and said it would all be simple if only it were not the case that, and it was so annoying that, Levent works at Anthropic. It was also said that if OpenAI posted after us, they would say that we deserved the Clay Prize, and that we were the “closest humans to the problem”. I declined both offers.
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
Wow, that's some VERY friendly communication. Besides, will the career of the person be ruined because of “Why would you ruin your career?” came out of his or her own mouth?
Thats how most people use LLMs? If I was back in my student days working in Navier-Stokes, I guess I would also punch at blow ups. The number of students doing this at the same time, posting open efforts to GitHub then retraining of the models. If there is solutions to the problem, it’s a real possibility that it was not a result of this effort?
Using a large amount of tokens is not a good augment that it’s not likely others have done the same. Good questions is the difference between $10 and $10M in token usage to solve a problem.
Strange times.
On the Navier–Stokes issue specifically, it has long been suspected that such a blow up would exist, and an AI telling you it indeed exists doesn’t contribute any new understanding to the field. And this problem seems like one that would be solved by humans anyways even if AI didn’t exist; accelerating the result by a few months/years using AI doesn’t mean much.
- OpenAI did related research around similar timeframe.
- Tristan claimed OpenAI offered a proposal that included dropping the Anthropic-affiliated co-author.
- Sebastian (a prominent OpenAI researcher involved) denied these claims.
- Tristan have no concrete evidence that OpenAI accessed their session.
- OpenAI's theory may hold up, but it will require long-term validation to confirm.
Separately, Terence Tao noted there is a low probability OpenAI actually solved the general regularity problem.
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models .
It’s true that Tristan has no concrete evidence OAI accessed their session. It’s impossible for him to have that without OAI’s say so.
Why are you framing things so pro OAI?
- He insisted that OpenAI initiated research based solely on rumors and never accessed their Codex sessions.
- He mistakenly believed Tristan and Levent were solving the same problem in Anthropic.
- proposed two option. (1) Tristan becoming the lead author to revise OpenAI’s work, or (2) OpenAI providing internal model to support and bridge their research.
- Sebastien insisted there was no intention to alter authorship. He was simply uncomfortable sharing OpenAI’s work,model with an Anthropic researcher. Additionally, He believed Levent’s credit seemed limited as their work focused on Euler.
- complained that negotiations with Tristan and Levent were difficult
[1] https://xcancel.com/SebastienBubeck/status/20973794116915163...
Its not about accessing the session, its whether the session ended up in the training run for the next version of the model
That's normal practice for these models and it seems like that was the case since OAI added a disclaimer
Obviously they can't causally prove it helped though since these models are incomprehensible
Terry also talks about it https://mathstodon.xyz/@tao/117234157753860650
- the OpenAI researchers claimed that they had "just told it to work on the problem" with little human input
- in fact, they had a whole team working on it
- and used, among other things, the work of third party human researchers to drive the work
- then threatened? a researcher who tried to go against theit planned narrative
Just from this document (which is of course only one side of the story) it really sounds like OpenAI was hoping to publish and say "we just told the model to try harder and it solved a Millennium problem!". Not great if true.
This part in particular was especially egregious:
> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”
"I don't want to live in a world where someone makes the world a better place, better than we do."
It's amazing how transparently OpenAI is running the standard silicon valley playbook.
but they probably won't and if it's happened on some obscure math research it's happening everyday everywhere else.
Fully local AI compute can't come fast enough, these guys have IP theft baked into their bones.
https://www.quantamagazine.org/computer-helps-prove-long-sou...
If I publish something, and disclose that I used AI for assistance, do I have to credit everyone who previously used the same AI to try the same problem? Because their prompts inevitably made it to the training data for my prompts?
Time will almost certainly reveal a lot more about the drama and the related ethics, but let's get excited about the actual breakthrough as well!
Tao mentioned "finding a configuration of water molecules that would collapse and shoot off to infinity", which would qualify IMHO.
OpenAI's board has fired Sam Altman. https://news.ycombinator.com/item?id=38309611
Apple sues OpenAI, accuses ex-employees of stealing trade secrets. https://news.ycombinator.com/item?id=48865019
While I support their argument - push for stronger data and privacy protections from OpenAI and similar - it is naive to believe we can have privacy while sending our data to third parties. It's clearly better to be safe than to be sorry here. Well, clearly better in terms of privacy. In terms of the maths gold rush, who can say what's better, that probably favours those taking more risk.
> We used several LLMs throughout: Anthropic’s Claude, OpenAI’s Codex, especially with GPT-5.6 Sol and, more recently, Astra. The latter was only used for writeups and auditing our arguments.
\nu d^2 u_i / dx_j dx_j - Viscosity
-1/\rho dp/dx_i - Pressure gradient
u_j du_i / dx_j - Advection. Kinda like momentum transfer from the motion of the fluid itself. Nonlinear, which makes the N-S equations hard to solve
du_i/dt - Rate of change of velocity. Note that this is in an Eulerian framework so it's not the acceleration of a packet of fluid, rather it's just the change in velocity at a particular location in space
Euler is when you omit some terms. Forcing is when you add some other terms to account for phenomena external to the fluid like gravity or flow through a porous medium like in the article.
And if we think this only applies to academic fields then we're doubly fooling ourselves. They do not have the ethics or incentives to be good stewards of the technology.
What I'm curious to know is whether this was a manual snooping, or automated farming that occurs for anything of value that happens in chats.
Ah yes, taking code from a person/company is fine if its deidentified!
But this (if it is true) is really abominable and a show of force from the techno-feudalists.
This should stop. What is possible does not mean it should be implemented.
Telling OpenAI that Anthropic has apparently solved an important problem but most likely that refers to him and he is using OpenAI models (not Anthropic's)?
And he wants to clarify that with OpenAI in advance? And get a pardon for Anthropic's likely but false press statements?
I dont get it.
[edited] needless to say, the behavior of the OpenAI employee is really despicable
https://xcancel.com/ElliotGlazer/status/2096298696438906934#
That seems like a valid reason to contact OpenAI.
$15m in tokens; but what about labor?
What about compressible fluids?
"Blowup for the Boussinesq equations with smooth forcing" (2026) https://cims.nyu.edu/~tristanb/boussinesq.pdf
"Extending the Córdoba–Martínez-Zoroa IPM Blow-Up to Uniformly Spacetime Smooth Forcing" (2026) https://cims.nyu.edu/~tristanb/ipm.pdf
"Finite Time Blowup for the Euler equation" (2026) https://cdn.openai.com/pdf/315b36cd-ec98-4023-8342-93345194e...
Lean proof: openai/NavierStokesAndEuler: https://github.com/openai/NavierStokesAndEuler
How to tree shake Lean4? (Edit: `lake shake`,)
[1] https://github.com/openai/NavierStokesAndEuler/tree/main/Nav...
I'm just wondering how much real input Buckmaster gave here that he thinks the proof is his. I guess at the end of the day OAI still wins if ChatGPT was used to prove this successfully.
Who still wants to use AI to solve cancer and other major problems?
(some drama from good ol' William)
Literally who cares who solved the problem just publish the results.
Academia was always politics first results second and I AM GLAD that LLMs are becoming superhuman at math. I like better theorems, not better politics.
It seems like OpenAI heard of the rumor and then scooped them because their internal model is better/they have more compute. OpenAI has NO obligation to mention Tristan nor Levent, because they DID NOT steal their data.
Imagine if OpenAI opened up a high frequency trading arm and suddenly stole all the prompts and research that other HfT firms are doing through OpenAI tools and start making bank based on that . Wouldn’t that be straight up insane?
Like, to me this looks like academic slap-fighting from Bubeck and Levent. People working at OpenAI are saying, "hey, we don't have that particular data in our models," others are saying, "we used a different approach to do it with Navier-Stokes" this feels like much ado about nothing.
Then in these comments I see some wild accusations.
If OpenAI is telling the truth (I don't really see a reason to lie here, if anything that sounds kind of like a dumb idea given the context), then they heard, "oh, shit, someone might be able to solve Navier-Stokes, don't we have some guys working on that? Give them 10,000 agents!" Then 88 hours later, out pops a similar solution. It's not like there's probably an infinity of ways to do this, the proof is probably similar.
Read this:
> Two proposals were offered to me. The first was that we post our Euler result, and that OpenAI post its Navier-Stokes result the next day. The second was that, after posting Euler, I alone write a paper presenting the Navier-Stokes result, acknowledging that an internal OpenAI model had resolved it. Sebastien twice asserted that he wanted Levent removed from authorship, and said it would all be simple if only it were not the case that, and it was so annoying that, Levent works at Anthropic. It was also said that if OpenAI posted after us, they would say that we deserved the Clay Prize, and that we were the “closest humans to the problem”. I declined both offers.
So, really, it sounds like academic slap-fighting nonsense and corporate bureaucracy. Literally, OpenAI's best move would have been to say, "ok, we're going to not say anything, do your thing" and let it happen. Ego and vanity got in the way.
Still, the stupid drama of this doesn't really do the results justice. There are maybe 1000 people on planet earth who are qualified to solve a problem like this. Even if the human "loosened the jar" a bit, that's astounding that their model was able to figure it the rest of the way out. Why are people dialed in to the human interest story here and not looking at the bigger picture!