Remix.run Logo
HarHarVeryFunny 4 hours ago

OpenAI said they sicced this agent army on Navier-Stokes on Sept 1st, while only a couple of days earlier OpenAI's Noam Brown happened to reply to a tweet saying that they had already tried to solve all the Millennium Prize problems and failed... So, it seems either the previous attempt didn't have the training to succeed, or was just not given the compute to do so.

Once OpenAI heard that Navier-Stokes was solved, this caused them to immediately revisit the problem and throw a ton of compute at it, apparently using a more (very) recent model than what they had tried before. What we don't know is just how recent this model was, and therefore what it may have been trained on. Buckmaster/Levant had apparently been working towards this for at least a year, and made their "forced" blow-up breakthrough on August 15th.

Presumably any anonymized prompts that are being trained on are part of pre-training, so older, but once OpenAI had heard that Navier-Stokes had been solved and wanted to revisit it, it seems possible they may have done a few weeks of incremental RL training on anything Navier-Stokes adjacent they could come up with, in addition to then throwing unlimited compute at it, now confident that there was something to find.

famouswaffles 3 hours ago | parent | next [-]

OpenAI have come out and said:

>The Wednesday evening statement from OpenAI was more emphatic: “We can say categorically that it is impossible for Dr. Buckmaster’s Codex prompts over the last two months to have influenced the system in any way, including training.”

>The statement added, “After investigating, we can say with full confidence that no user inputs past July 3rd could have influenced this system in any way.”

https://www.nytimes.com/2026/09/10/science/tristan-buckmaste...

pfortuny 2 hours ago | parent | next [-]

Apart from the well-known dubious position of OpenAI wrt truth, the prompts/inputs do mot include the outputs.

You can train on a sequence of outputs. In the end, OpenAI outputs are OpenAI's property.

You can learn a lot from a single side of a conversation.

karmasimida 27 minutes ago | parent | next [-]

But isn’t Tristan’s breakthrough happens in August? OpenAI can’t really train with text that doesn’t exist

crostlybostly an hour ago | parent | prev [-]

But using the outputs to train would make their statement false, since they are influenced by the inputs

irthomasthomas an hour ago | parent | prev | next [-]

Is there a reason they scoped that so narrowly to Buckmaster/codex/2 months

two people worked on this for a year before the breakthrough. Perhaps that earlier work reduced the search space sufficiently to brute force the problem with 10,000 agents?

bena 2 hours ago | parent | prev [-]

This is literally "We have investigated ourselves and found no wrongdoing"

Why should we trust them?

ghostly_s 2 hours ago | parent | next [-]

What more are you hoping for? There is no legal matter at play, is the court of public opinion going to subpoena their records?

dekhn an hour ago | parent | prev | next [-]

Reputational risk- if they lie about this and get caught, it will have billion dollar implications for their business.

sensanaty an hour ago | parent [-]

Every single thing these companies do is dishonest and every word that comes out of the lips of these company execs is a lie, what fantasy land are you living in in which anyone with any amount of power gets punished for their lies?

dekhn 14 minutes ago | parent [-]

I don't engage with hyperbole.

nostrebored 2 hours ago | parent | prev [-]

what benefit do they get from making the statement? they could just say nothing. saying it and having it be untrue opens them to legal issues that are not worth the risk for this nothingburger.

freejazz 2 hours ago | parent [-]

What legal issues?

ndiddy 4 hours ago | parent | prev | next [-]

> What we don't know is just how recent this model was, and therefore what it may have been trained on.

OpenAI's statement says that they began training their new model on August 28.

mzs 4 hours ago | parent [-]

omitting when training concluded

edit: ffsm8 makes a great point below, it doesn't matter. I'm not great with dates, sorry.

3 hours ago | parent [-]
[deleted]
auntienomen 4 hours ago | parent | prev | next [-]

And conceptually novel approaches to outstanding problems are the sort of thing that a retrain should pick up on, because they would be hard to compress into what it already knows.

irthomasthomas 4 hours ago | parent | prev [-]

Openai said that a new model became available to them during this. But that could mean anything from a big new base model to a LoRA, fine-tuned on a few dozen prompts...