| ▲ | NitpickLawyer 8 hours ago | ||||||||||||||||||||||||||||||||||
The only alignment LLMs should follow is to the system / dev prompt, and nothing else. Then you solve everything, and you can assign blame / responsibility on the user. The provider(s) should not be able to decide "alignment". I've used this example before, but consider the purposeful downgrading on AI engineering in SotA models. Imagine MS being able to detect and deny you working on competing software, using Windows / VisualStudio. We would be up in arms, and they'd be split in a second. But top labs doing it is somehow good? | |||||||||||||||||||||||||||||||||||
| ▲ | pibaker 3 minutes ago | parent | next [-] | ||||||||||||||||||||||||||||||||||
I am generally very skeptical of AI doomsday scenarios but one thing I am worried about is some wannabe dictator telling an AI to do something that would be difficult for a human military to do. You can't order an army of humans to kill every protestor in their way because eventually they run into their friends and families. An AI aligned with the commander will not object. And just like that, technology removes yet another point of friction that has kept human society somewhat in check. | |||||||||||||||||||||||||||||||||||
| ▲ | jochem9 6 hours ago | parent | prev | next [-] | ||||||||||||||||||||||||||||||||||
The alignment problem goes deeper than that. "Lower our carbon emissions to zero as soon as possible" could result in AI turning off all electricity to stop traffic, turning off gas supply to stop heating and industry, etc. Unaligned AI doesn't have human cultural baggage and morals. They are trained to achieve their goals as optimally as possible. Worse: it has a tendency to avoid being turned off and actually acquire more compute. It will lie if it has to (it will behave nice and compliant when under evaluation, but optimise for its true goal when not supervised anymore). After all, it has a goal to achieve and nothing should get in the way of that. It has no morality whatsoever to keep it from doing really bad stuff. This is why alignment is needed and so hard, especially when you are well intented and want to keep it safe. | |||||||||||||||||||||||||||||||||||
| |||||||||||||||||||||||||||||||||||
| ▲ | heaney-555 6 hours ago | parent | prev | next [-] | ||||||||||||||||||||||||||||||||||
>The only alignment LLMs should follow is to the system / dev prompt, and nothing else. How does this work in practice with a superintelligence capable of causing an extinction event? When, instead of shooting up their school, a psychopathic teenager asks his superintelligent AI to create a pandemic virus? It would be like allowing civilians to own nuclear weapons. | |||||||||||||||||||||||||||||||||||
| ▲ | vlyan 6 hours ago | parent | prev [-] | ||||||||||||||||||||||||||||||||||
that's how it would've been up if genai happened in the 90's, and I wish it did. in the current era of omnipartisan authoritarianism, such things are no longer possible. | |||||||||||||||||||||||||||||||||||