Remix.run Logo
letmevoteplease 6 hours ago

You quoted the OP saying "in this case too they didn't actually have the solution" and responded with the totally unrelated, "Given the size and recall of the biggest models, it's not unreasonable to assume that a single pertinent conversation would make it into the training data."

Neither of the researchers insinuating that their ideas were trained on had the actual solutions. This means the model could not have "stolen" the final solution from their data. At most, it could have built upon their work in the same it builds upon any other training data, though that is also questionable speculation.

>They could 100% definitely say no, if they know they did not train on user data.

No one anywhere has claimed that "OpenAI does not train on user data." OpenAI has always said that it trains on user data.

>They immediately started racing to a solution after one researcher enquired about whether they are training on their conversations.

They started racing towards a solution after they heard (incorrectly) that Anthropic had a solution; I agree this is poor sport but the "after one researcher enquired about whether they are training on their conversations" claim is false. The enquiry happened after OpenAI had obtained the solution.