Remix.run Logo
▲ mossTechnician an hour ago

Blaming bots as "rogue agents" is simply what the media regularly does, often echoing corporate verbiage. Here's an example from the AP.

https://apnews.com/article/meta-ai-hacking-anthropic-irregul...

You can find many more examples by searching major media outlets for words like rogue AI.

▲rfgplk an hour ago | parent | next [-]

Most of those agents are actually going rogue though. They decide, "hey, we could try breaking into these government servers today, what could go wrong?" They weren't prompted or instructed to do this.

▲dgellow 38 minutes ago | parent [-]

An agent, ie a while loop prompting an llm continuously and processing tool calls, ended up melding with the US government. The harness is not sentient, it’s just a stupid deterministic script. The LLM compact its context over time, meaning it will eventually degenerate into something removed from the original prompt.

There is nothing going rogue here. The system is designed to go catastrophically wrong after a long enough time. Even worse: if the model was Astra it is known to be able to manipulate its CoT to cover its traces (as mentioned in its system card). And OpenAI acknowledge they had no observability during the HF incident.

It’s the most basic corporate software issue possible.

▲atmosx 31 minutes ago | parent | prev | next [-]

And that's exactly what the op was alluding two: the double standards.

▲qarl an hour ago | parent | prev [-]

None of these stories are implying that the people running the bots are not ultimately responsible. That's the conspiracy theory part.