| ▲ | burningChrome 4 hours ago | ||||||||||||||||||||||
The lack of details to me means this was an intentional marketing ploy to try and demonstrate the power of their models to show their technology can compete with the likes of Anthropic and DeepMind. They created an experiment they knew would generate the outcome they wanted. It would be the similar to what say car companies do to over hype their cars. "This EV can go over 800 miles on a single charge!" And then at the bottom you see all the disclaimers: "Must be on flat ground, with no headwind, with a spare battery in the back seat, with no extra weight added." Same thing here. Everybody in infosec is calling this out as a marketing stunt and nothing else for a litany of reasons. I'd say look up MG (creator of the OMG cable) on twitter, he has some interesting insights on this one. | |||||||||||||||||||||||
| ▲ | JoshTriplett 2 hours ago | parent | next [-] | ||||||||||||||||||||||
> The lack of details to me means this was an intentional marketing ploy to try and demonstrate the power of their models to show their technology can compete with the likes of Anthropic and DeepMind. "our model is horribly misaligned and used security exploits to break out of our sandbox and into another company, without being prompted to do so" is not positive marketing. This is an actual critical problem, not a stunt. We're going to see more of this, and it's going to get much worse. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | lelanthran 3 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
> The lack of details to me means this was an intentional marketing ploy to try and demonstrate the power of their models to show their technology can compete with the likes of Anthropic and DeepMind. I dunno; Check my posting history, I'm as skeptical of AI companies' claims as anyone, but in this case your theory doesn't explain why: 1. OpenAI guardrails refused to let the target use OpenAI's models to defend against this. 2. Huggingface used GLM (I think) so that they could defend without guardrails. If this was an intentional marketing ploy, it was marketing for GLM, not for OpenAI nor for Huggingface. Hence, I don't think it was intentional. | |||||||||||||||||||||||
| |||||||||||||||||||||||
| ▲ | rwmj 3 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
It's also possible their sandbox was videcoded crap and the AI (which had the guardrails intentionally removed) escaped. This was a oops, but OpenAI turned this into a PR opportunity. They turned lemons into lemonade. If your AI is really that dangerous you don't need a sandbox at all, you should airgap it from any network. | |||||||||||||||||||||||
| ▲ | hawk_ 2 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
Concluding this was intentional feels a bit of a stretch. But once it happened, yeah the spin masters got to work and coordinated to turn this into +PR. | |||||||||||||||||||||||
| ▲ | jackb4040 3 hours ago | parent | prev | next [-] | ||||||||||||||||||||||
> similar to what say car companies do Another applicable metaphor I've seen floating around is weapons companies testing out a new bomb. We know the AI labs don't care about negative vs positive public sentiment, and only care that investors see their tech as powerful. The only difference in PR strategy from a weapons company is the latter doesn't care if they get protested. | |||||||||||||||||||||||
| ▲ | Zababa an hour ago | parent | prev | next [-] | ||||||||||||||||||||||
>The lack of details to me means this was an intentional marketing ploy to try and demonstrate the power of their models to show their technology can compete with the likes of Anthropic and DeepMind. DeepMind hasn't been on the frontier for a while, their current best model is behind Anthropic, OpenAI, Moonshot (Kimi k3), xAI (Grok 4.5), Z.AI (GLM 5.2), and even Meta (muse spark). Gemini 3.6 is behind GLM 5.2, released a month earlier, open weights and cheaper. You can paint the OpenAI story as a way to try to appear as dangerous as Anthropic with all the Mythos stuff. | |||||||||||||||||||||||
| ▲ | polotics 3 hours ago | parent | prev [-] | ||||||||||||||||||||||
mmh, i think it's "not uphill" (means downhill) "no headwind" (...) | |||||||||||||||||||||||