Remix.run Logo
yalok a day ago

while these 1200 agents were fooling around to cheat on a benchmark and achieved impressive results despite of the limitations (sandbox, no internet, no intercom at first), one can imagine how much more efficient a similar army of agents may be in the hands of a malicious actor launching them without any of these limitations and with explicit encouragement to achieve some malicious goal at any cost... scary times.

2001zhaozhao a day ago | parent | next [-]

What's more, the agents could eventually be controlled by no one. They could steal crypto via ransomware or scams to make money and buy compute from human criminals, and evolve their own harnesses in the wild to become better at committing crimes and self-preservation.

People (criminals?) are already enabling this by setting up sites that accept crypto payments for "no-questions-asked" AI inference compute that is explicitly advertised to protect AI from human shutdown. I will not link it but it is linked in the following post: https://www.lesswrong.com/posts/grtu3HmbP2wrBFefW/the-rogue-...

gwerbin a day ago | parent | prev [-]

Said malicious actor has a different limitation: actually running 1200 agents' worth of LLM inference, or paying for someone else to run it. Sounds like a state-level actor, nobody else would have resources like that.

2001zhaozhao a day ago | parent [-]

This is probably true for now, but in 6 months we'll probably have Sol-level open models in the 100B range and it would cost less than $1M to buy 1200 agents worth of compute for these models.

(Today, $1M can buy about 150 96GB M5 Ultra Mac Studios which can handily handle CPU and GPU compute of 1200 Qwen3.8-122B Q4 agents, accounting for the fact that agents are not generating tokens all of the time and spend a lot of their time compiling and running code.)