Remix.run Logo
keremk 2 hours ago

Occam's Razor says: "They heard this problem is solved or about to be solved amongst the rest of the other problems. They prioritized this and put substantial compute with their newest model and solved it." I know everyone loves juicy rumors, theories etc. but honestly that is the simplest and most plausible explanation given the state of AI improvement now. Obviously spending 15 million on a problem is not a slam dunk decision even for a company like OpenAI but if it has a significantly high chance of solving it and their competitor will be claiming they solved it, then it raises the stakes and they go after it. In fact this is the most rational and also curiosity-driven thing to do and totally what I would have expected from any frontier lab. Of course if one wants to prove their confirmation biases that they train on sessions or be able to identify individual users, the non-zero chance of that being also another explanation is attractive enough to wet their appetites.

davesque an hour ago | parent | next [-]

Isn't it also a simple idea that a model designed to recall relevant information from its training data, which is also known to have been trained on data from user transcripts, would, in fact, reproduce directly relevant work by leading experts in the field? Seems like Occam's razor would apply to that situation as well.

We know that LLMs are trained to recall relevant info. We know AI vendors are using user transcripts to train models. Two plus two equals four, right? I mean, an LLM that failed to recall the transcripts of those researchers would be a bad model.

PowerElectronix 2 hours ago | parent | prev | next [-]

Funny how they did not solve any of the other problems, just the one where there was already solutions to the NS with some restrictions in their chats, and their solution seems to derive from those.

jojva an hour ago | parent | next [-]

They explain in the article that they eventually redirected all their resources towards NS.

consp an hour ago | parent [-]

So instead of dictating research to unsolved, or largely unsolved, problems, we are now as "a society" directing compute power towards sniping research outcomes.

I thought this was an ebay thing for people with too much free money, but it seems a bit larger.

OtherShrezzing an hour ago | parent | prev [-]

Considering how language models work, you'd expect two distinct conversations on the same mathematical problem to have enormous crossover.

l5870uoo9y 2 hours ago | parent | prev | next [-]

Given that the solution took a somewhat “unusual” approach, I find it even more unlikely that an AI model would have come up with this on its own.

pietz an hour ago | parent [-]

Isn't that *exactly* the type of solution you'd expect from AI?

Move 37 comes to mind.

iLoveOncall an hour ago | parent [-]

Only if you understand nothing about the difference between LLMs and AlphaGo.

npiano 2 hours ago | parent | prev | next [-]

Why is that simple or plausible? Why is simpler or more plausible than lifting an almost-finished solution from a researcher's account?

ghshephard 2 hours ago | parent | next [-]

Because they solved different problems, and where there was overlap, the solutions look different?

karmasimida 2 hours ago | parent | prev [-]

Because it is not finished. Their follow up claim is that OpenAI’s approach looks like another proof they had been working on the side, but hasn’t published yet

20k 2 hours ago | parent | next [-]

OpenAI have admitted their new model they used was trained on prompts at around the time that researcher was working on it, so it seems self evident that it was used as part of the millennium solution

spwa4 an hour ago | parent | prev [-]

1) Because AI models are 1000x better about following a problem to its conclusion than coming up with a genuinely new idea.

2) Because if we accept the facts ChatGPT only came up with its "new" idea after being told exactly what the new idea was by a mathematician (OpenAI doesn't dispute this btw). And OpenAIs story comes down to the usual "We didn't look at it, trust me bro", which is made more hard to believe because OpenAI only started their efforts after receiving news of what the researcher was doing.

Oh and OpenAI emphasizes that part of the researcher's progress was made ... on OpenAI.

3) And, probably, the researchers were likely stopped by token limits, and that's the only reason they were slower than OpenAI themselves, which is very, very unfair.

4) OpenAI's story "smells" (like so many AI stories lately). Supposedly the company's team asked ChatGPT about solving millennium problems, and out of all millennium problems it just happens to pick the one where a solution can be found in its chat logs?

5) Yet again it would be in good taste for these AI companies to just give this to the researchers (no shortage of difficult unsolved math problems, so if AI can solve them all, just find another one). But instead, yet again they're fighting about it.

6) OpenAI admits they only went after this problem, with a team, no less, after finding out which researcher went after what problem, because of how they thought it would affect ChatGPT's PR. They are demonstrating, in other words, their willingness to destroy human researcher's reputation for PR wins.

That's getting close to big tobacco level morals right there.

7) If anyone wants to verify how much OpenAI cares about the truth, just ask on ChatGPT about the copyright lawsuit outcome and how it applies to OpenAI.

You'll get EXACTLY the sort of responses you get from Qwen about Tiananmen ("we didn't do it, you have no data, everyone's lying and if we did do it, it was perfectly reasonable because " style argument. Try it)

ozgung an hour ago | parent | prev | next [-]

I agree with Occam.

That 100x step up from using 100 agents for Euler to 100_000 agents for Navier-Stokes, in a single day seems a bit sus.

I also think this is not ethical behavior. This is at least academic dishonesty, kind of a plagiarism or intellectual theft.

That’s why they wanted to credit Tristan and to give $1M award to him. But again, they acted unethically in that process as well. They wanted him to remove Levent (Anthropic affiliation) as co-author and threatened Tristan to “end his career”. Their behavior is actually telling, their work was not completely independent from Tristan&Levent’s unpublished work.

Bluestein an hour ago | parent | prev | next [-]

There's also an elephant here in this room: Quite "coincidental" one of the coauthors worked for the competition. Of course GPT would know this fact. And OAI has every incentive to particularly monitor those accounts.-

alangibson 2 hours ago | parent | prev | next [-]

This is not the simplest explanation.

The most direct line from problem to proof is OpenAI building off of conversations the mathematicians had with their AI.

Rebuff5007 an hour ago | parent | prev [-]

[dead]