Remix.run Logo
efnx 5 hours ago

I heard a rumor (on instagram, so YMMV) that the professor who was closest to solving this problem had only weeks ago used Codex, which had slurped up all his notes on the subject. Now OpenAI's agents solve the problem. If it's true that seems like quite a coincidence.

1121redblackgo 4 hours ago | parent | next [-]

See other thread, but yeah that's the general ballpark of the situation.

scuppernong 4 hours ago | parent | prev | next [-]

this is not a rumor (the allegation, anyway), it's reported in the new york times

jcranmer 4 hours ago | parent [-]

Newspapers are not above printing rumors.

See, e.g., Barak Ravid regularly reporting in Axios the impending ceasefire negotiation progress in the Iran War, which largely have failed to come to pass.

efnx 4 hours ago | parent | prev | next [-]

I don't understand why I'm getting downvoted, I'm not posting an opinion. Coincidences happen. So does foul play. No judgement call here.

dooglius 4 hours ago | parent [-]

There have been several threads and developments on this over the past few days, including statements from the primary subjects involved. Third-hand instagram comments are not really the best source to be bringing in.

efnx 4 hours ago | parent [-]

Everybody comes into information in different ways. There were no comments here about this specific aspect of the story - which is definitely interesting!

bethekidyouwant 4 hours ago | parent | prev [-]

How could they possibly included in the previous training run which takes months to complete..

mswphd 4 hours ago | parent | next [-]

I won't take a side in things, but OpenAI stated the model they used here started training August 28th. Note that "training" here might mean "post-training with RLHF an Astra base model" or something. but training had only started a little over a week earlier.

jazzyjackson an hour ago | parent | prev | next [-]

Has it not been the usual process to snapshot a model to use for inference while continuing to run the training process? I guess you can’t add to the training corpus once you begin? Just trying to make sense of whether training begins or ends as rigidly as you suggest.

s900mhz 4 hours ago | parent | prev | next [-]

IMO It’s not about being trained on the data, it’s more like what do the agents have access to during inference? Can they grep customer transcripts/logs?

bethekidyouwant an hour ago | parent [-]

You’re saying that when they we’re trying to solve this theorem they also shoved in its context somebody else’s chat logs? Bruh.

metanonsense 4 hours ago | parent | prev [-]

Maybe the boundaries of the memory subsystem are a bit fuzzy.