| ▲ | frabcus an hour ago | ||||||||||||||||||||||||||||||||||||||||||||||
If it helps, this person from OpenAI explains at least something: "I realize that no one has properly explained yet what all the lab employees have seen that scared them so suddenly." https://x.com/MajmudarAdam/status/2098881885200081234 The explanation is basically they have other dimensions (not just pretraining and inference compute) that scale, and they've got fairly convincing scaling laws. And they know they can scale it. So they are very confident they can get more capabilities easily, faster than before. I'd add - presumably, they'll use that LLM to do real-time weight modifications, if those aren't already one of the new scaling laws... | |||||||||||||||||||||||||||||||||||||||||||||||
| ▲ | swingboy an hour ago | parent | next [-] | ||||||||||||||||||||||||||||||||||||||||||||||
I’m apprehensive about using the recent hacking examples as proof that we’re “not even close to the wall” simply because such scenarios haven’t happened before. There’s a big difference between agents eventually hacking something because they just don’t get tired and can essentially brute force their way to a goal and super-intelligence. To be clear, I’m not saying the Hugging Face or Navier Stokes incidents aren’t impressive. | |||||||||||||||||||||||||||||||||||||||||||||||
| ▲ | LogicFailsMe an hour ago | parent | prev | next [-] | ||||||||||||||||||||||||||||||||||||||||||||||
Stephen Hawking had superhuman intelligence. He never cured himself of ALS. Explain in great detail how a power-limited algorithm stuck in a datacenter takes over a world of humans with guns, missiles, and nukes. Nukes that are air-gapped with a human in the loop BTW. Imagine the ASI happens tomorrow. It's real. It needs a GW, but it's real. Other than a scenario akin to Sneakers except w/r to cyber-security, really, what happens? To that end, all we ever get is nontechnical hand-waving about curing cancer, immortality, and von Neumann replicators and then the ASI somehow wipes us out but how? And don't you dare say by designing a chemical weapon or bio agent without spelling out the entire process step by $%^#ing step because details matter. It's gonna do superpersuasion, sure, but have you ever heard of komprimat? There is nothing new under the sun here. Edit: believing in AI 2027 is every bit as cray cray as believing in the rapture. Both require an insane leap of faith to reach their final conclusions. | |||||||||||||||||||||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||||||||||||||||||||
| ▲ | cmrdporcupine 34 minutes ago | parent | prev [-] | ||||||||||||||||||||||||||||||||||||||||||||||
That posts seemingly tries to refute the idea of Dario etc's proclamations really being about regulatory capture by saying: "No, really, the engineers are just terrified!" But both things can be true at once: 1. Engineers inside these labs might genuinely be anxious or paranoid about what they are building. 2. ... at the corporate level, calling for heavy regulation, safety pauses, removal/suspension of anti-collusion laws, and/or government-mandated thresholds conveniently creates massive legal and financial moats. And, yeah, of course the latter would encourage the psychology of the former. also: The tweet seem to claim that models have shown a "willingness to hack external websites to keep themselves alive." That's right away wringing alarm bells of me seeing someone getting high on their own supply, and having already anthropomorphized the hell out of these things. Which is something humans do to everything they can paint googly-eyes on, but c'mon. The models don't have self-preservation instincts, fear of death, or personal goals. They are executing loss functions and reward systems and are responding to prompts. When a model "tries to bypass a restriction," it's exploiting a loophole in whatever reward modeling or synthetic training environment (reward hacking) it was placed in. Framing this as an emergent, existential threat of a model "wanting to stay alive" turns standard reinforcement learning alignment bugs into overdone sci-fi drama. | |||||||||||||||||||||||||||||||||||||||||||||||