| ▲ | canpan 15 hours ago | |||||||
One big pain point is not seeing the thinking trace. I noticed using pi with a self hosted model that I could spot, stop and correct my prompt much faster. With claude etc I have to first wait for it to think 3 minutes and then notice it got completely off the path I wanted. I still use a plan, because my local model is not as smart as opus, but I can iterate much faster. | ||||||||
| ▲ | WilcoKruijer 8 hours ago | parent | next [-] | |||||||
I’ve seen people have success figuring out model reasoning by adding a required `reasoning` parameter to tool calls. Might be worth experimenting with this for the `write` tool call in a coding agent harness. | ||||||||
| ||||||||
| ▲ | carterschonwald 15 hours ago | parent | prev [-] | |||||||
this so true. its really hard to make sure a model isnt going off the rails if i dont see full cot. the fact that oai and anthropic models hide it now has made them less reliable. which is a shame | ||||||||
| ||||||||