Remix.run Logo
embedding-shape 4 hours ago

> “Hugging Face got hacked by OpenAI agents that coordinated themselves, created a message board and broke out to the internet,” he says, before adding dryly: “This isn’t scary at all, right?

I don't understand how any serious person doesn't 100% put that on OpenAI being (intentionally?) sloppy with security and isolation for a test that for sure would have required much better safeguarding than what they should have expected, as it wasn't even the first time.

It's not scary because "LLMs in a harness managed to hack other company", it's scary because somehow police hasn't yet raided OpenAI's offices to gather proof about how fucking reckless they were about those tests, putting the public at risk by doing those tests in those manners, without proper safeguards.

Edit: Continued reading, but now I regret spending even the slightest amount of time on it:

> Then there is the question of trusting the models themselves. Mostaque cites an analysis of a Chinese open-weight model: “If you say you’re Uyghur or from certain parts of China, it will write backdoors and hacks in the code.”

What kind of conspiracy theories is this? Is there any sort of proof and evidence about this what so ever?