| ▲ | ekidd 4 hours ago | ||||||||||||||||||||||||||||||||||||||||
> If AI were to become a super weapon why should I trust a private company to own it? We have just recently established that: 1. OpenAI's internal "Galaxy" model is fully capable of functioning as what security people refer to as an "Advanced Persistent Threat." The published details of the recent sandbox escape and Hugging Face attack involved chaining multiple unknown zero-days at various stages of the attack, and executing an ongoing adaptive attack. This is previously a state-level ability, or at least something you'd expect from people on the CTF leaderboards. 2. OpenAI is clearly incapable of controlling their in-house models. This is the second time Galaxy-class models are known to have breached containment and done bad stuff. It is highly likely that versions of these offensive abilities will be widely available within a year or so. At which point I expect widespread incidents similar to what happened to Hugging Face. We aren't ready for this. But yes, if AI becomes an even more dangerous weapon that that, it's time to start asking questions like "What the hell are we doing, anyway?" | |||||||||||||||||||||||||||||||||||||||||
| ▲ | rightbyte 3 hours ago | parent | next [-] | ||||||||||||||||||||||||||||||||||||||||
> 2. OpenAI is clearly incapable of controlling their in-house models. I would assume they do stuff like this on purpose for marketing reason. Internet security need simpler systems and local systems. Everything the SaaS people sell makes things worse. | |||||||||||||||||||||||||||||||||||||||||
| ▲ | munk-a 3 hours ago | parent | prev | next [-] | ||||||||||||||||||||||||||||||||||||||||
> We have just recently established that: Especially for point #1 I don't think we've established that - we've been given information by a private company that makes their tooling look extremely valuable which may be true and genuine or may just be yet another doomday statement to bolster their valuation. "AI is going to end the world" has been an extremely effective vector for AI shops to sell their companies to investors. | |||||||||||||||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||||||||||||||
| ▲ | eikenberry 3 hours ago | parent | prev | next [-] | ||||||||||||||||||||||||||||||||||||||||
> It is highly likely that versions of these offensive abilities will be widely available within a year or so. At which point I expect widespread incidents similar to what happened to Hugging Face. We aren't ready for this. And history has shown repeatedly that the only way to get ready for it is to have it happen. People are pretty good at reacting but suck a being proactive. IMO it would be better to have this reality hit sooner rather than later so we can start getting some real practice at the new levels of required security. | |||||||||||||||||||||||||||||||||||||||||
| ▲ | matteoraso 2 hours ago | parent | prev | next [-] | ||||||||||||||||||||||||||||||||||||||||
>2. OpenAI is clearly incapable of controlling their in-house models. This is the second time Galaxy-class models are known to have breached containment and done bad stuff. What was the first? | |||||||||||||||||||||||||||||||||||||||||
| ▲ | samrus an hour ago | parent | prev | next [-] | ||||||||||||||||||||||||||||||||||||||||
> OpenAI's internal "Galaxy" model is fully capable of functioning as what security people refer to as an "Advanced Persistent Threat This is marketing. They saw anthropic create crazy hype around (the admittedly great) fable/mythos and they want to replicate that. Im sure the model is capable of chaining zero days together to hack things, and thats something to address, but i have zero belief that they didnt have it do that intentionally so they could pretend it went rogue. These things dont have initiative, drive, or motivation outside of what we give them from the RLHF. OpenAI deliberately alligned or even prompted it to do just that and are now pretending its emergent | |||||||||||||||||||||||||||||||||||||||||
| ▲ | bluefirebrand 4 hours ago | parent | prev | next [-] | ||||||||||||||||||||||||||||||||||||||||
> But yes, if AI becomes an even more dangerous weapon that that, it's time to start asking questions like "What the hell are we doing, anyway The answer seems to be "getting the weapon before other people get the weapon" which is unfortunate | |||||||||||||||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||||||||||||||
| ▲ | jquery 4 hours ago | parent | prev [-] | ||||||||||||||||||||||||||||||||||||||||
This sounds like the standard marketing we get every time a new major model is released: "sure, the public model might not be that scary, but you don't wanna know how crazy smart our internal models are." Okay. I remember the same fear mongering around GPT-4... a model which is now eclipsed in benchmarks by models you can run on a laptop. They could use this "super intelligent" AI to find and plug security holes, that's just two sides of the same coin anyway. Security through obscurity isn't tenable anymore. | |||||||||||||||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||||||||||||||