| ▲ | gck1 4 hours ago | ||||||||||||||||||||||
They had a model escape in April, roughly the same time when they were fearmongering about Mythos and how Anthropic should be the sole keyholder of cybersecurity capabilities, and it only occured to them to look inside logs when they saw someone else winning in their own game. What, Anthropic didn't know model could escape sandbox without OpenAI reporting it? | |||||||||||||||||||||||
| ▲ | skeptic_ai 4 hours ago | parent [-] | ||||||||||||||||||||||
Yeah, the company that only says “safety” every other 3 words, they don’t even think to have a fake decoy internet to alert them mechanically about any internet access limitation bypasses? See more https://news.ycombinator.com/item?id=49117555 Also simonw stance on this i’d say it’s at least concerning… seems like he is here to keep a good image (or better said less bad) of anthropic. | |||||||||||||||||||||||
| |||||||||||||||||||||||