| ▲ | fermuch 7 hours ago | |||||||
xhigh tells it to overthink and re check everything. Low tells it to only do the minimum thinking necessary. I would suggest to give qwen medium which doesn't inject any thinking directives into it and also to give as much context as you can, ideally around 500k tokens or even 1M if you can. Big complex tasks like these make the model hit the compaction trigger a lot and they end up re thinking the same thing several times in my experience. | ||||||||
| ▲ | kennywinker 5 hours ago | parent [-] | |||||||
Doesn’t it max out its context at like 256k? | ||||||||
| ||||||||