| ▲ | arjie 4 hours ago | |
Is it actually entirely a prompt-based information? I’d assume that some of it is the harness part of the agent setting reasoning token budget and compacting reasoning etc. In that case, the agent will respond incorrectly because it has no visibility into what reasoning mode it’s in. | ||
| ▲ | willy_k 4 hours ago | parent [-] | |
IIRC responding to effort level settings appropriately is part of the (post)-training. In that case it could be considered another instance of the Bitter Lesson. Uplifting. | ||