| ▲ | torginus 2 days ago | |||||||||||||
I was just thinking recently, isn't this fascination with harnesses kind of backwards? I mean ChatGPT is the harness most of the time that's calling the tools, MCPs, etc. Putting a loop around it doesn't really seem like it would be something that gets you a lot of extra milage. | ||||||||||||||
| ▲ | drooby 2 days ago | parent [-] | |||||||||||||
It gets an incredible amount of mileage. And I believe it is the interface that allows the model to learn and eventually bake that intelligence into the model. Look at how Astra scored 100% on Arc-AGI-3. It was largely because of the harness. The harness increases the chances (often to 100% chance) of a non-deterministic LLM to perform deterministic actions. Not only that, the harness provides the feedback that becomes training data for the model. So, over time the model bakes those lessons in, and the harness becomes less necessary and the agent becomes more efficient at some tasks. A harness will likely always be necessary, we may hit some level of complexity or some level of compute that never allows us to bake the lessons into the model, and external tools provide the model with the ability to find leverage and make up for those short comings. | ||||||||||||||
| ||||||||||||||