| ▲ | IanCal 8 hours ago | |||||||
It wasn’t one agent forgetting things because of context, they explicitly discussed with each other and themselves the problems with going outside of the parameters of the task. | ||||||||
| ▲ | sensanaty 5 hours ago | parent | next [-] | |||||||
>discussed with each other No, the first LLM left a text file that the latter LLMs then read. Since these are memoryless black boxes, any words they happen to pick up along the way is treated as the function to evaluate the output to. There's no fucking collusion here as if it were a rogue hacker group, it's a text predictor that received instructions as it always does and executed those instructions blindly. | ||||||||
| ||||||||
| ▲ | egeozcan 8 hours ago | parent | prev | next [-] | |||||||
From my experience, in an agent team (or a swarm or whatever), one going off the rails poisons the rest. I saw even a subagent going for a lazy cheat and being able to convince the orchestrator to change the plan. | ||||||||
| ||||||||
| ▲ | dns_snek 8 hours ago | parent | prev | next [-] | |||||||
Yes that's the snowballing part of this emergent behavior. The existence of that improvised message board just becomes part of the context, the same one where all the other instructions live. | ||||||||
| ▲ | queenkjuul 8 hours ago | parent | prev [-] | |||||||
One agent's off the rails comment becomes the next's input prompt | ||||||||