| ▲ | fancyfredbot 5 hours ago | |||||||||||||
When You See One Cockroach, There's Probably More. This can and should erode our trust in every single proof OpenAI published. The model is clearly faliable despite the lean proof, and clearly the output was't actually checked properly before release. Once these proofs are peer reviewed and published in a journal we might be able to trust them again but until then they are just slop, sadly. I think OpenAI actually did the right thing by sharing everything with the whole community right now but I also hope that some significant credit will now go to the reviewers who confirm these 'proofs" actually work. | ||||||||||||||
| ▲ | vlovich123 5 hours ago | parent [-] | |||||||||||||
Fantastic. Let’s also apply this reasoning to math papers written by humans too? Humans produce flawed Lean proofs. Indeed LLMs were successful at finding and fixing many issues in the “core” standard library if I recall correctly. Humans regularly produce flawed papers and have minor issues require fixing. And when it happens it often isn’t as prompt and clear as this. | ||||||||||||||
| ||||||||||||||