| ▲ | XenophileJKO 12 hours ago | |||||||
I probably have an unpopular viewpoint. I think the hacks really showed that the alignment is kind of mostly working. There is still a lot of work to do and the agents did some damage, but the blast radius was pretty small and they could have done much more damage if alignment didn't hold up as well. Edit: Though I do want to be clear it does also show how a maligned model probably can do some serious damage. | ||||||||
| ▲ | sajithdilshan 10 hours ago | parent [-] | |||||||
> There is still a lot of work to do and the agents did some damage If it was actually damage and disruptive shouldn’t police or the respective authority do an investigation on that? Because that’s usually what happens when a damage or harm is done by someone. That way we would actually be able to see a third party independent investigation on how dangerous the agents are. Right now we’re just blindly trusting whatever AI companies are saying. What is they aimed the agents on purpose to do those tasks to push their narrative and push for regulation | ||||||||
| ||||||||