| ▲ | anon373839 5 hours ago | |
> Qwen 3.6 35B A3B on medium thinking mode Qwen 3.6 doesn’t have configurable reasoning effort, does it? | ||
| ▲ | dofm 4 hours ago | parent [-] | |
Hm — brain jumped tracks a bit there at nearly 4am. I'm talking about budget — I mean limiting it to 2048 tokens. … for one or other of the models I tested at the same time, in llama-server, there is a dropdown that offered options (unlimited, max, medium which was 2048) (I've tested so many of these things now that they are beginning to blur.) I thought that was llama-server with Qwen 35B, just checked and it's not. Nor is it Gemma 4 26B. Perhaps it was Ternary Bonsai which I tested again and deleted earlier. Anyway I took to clipping Qwen 3.6 35B at 2048 tokens reasoning in LM Studio and elsewhere, and it did OK at that (because it often loops like mad on an ambiguous prompt if not curtailed). FWIW I just rechecked outputs and I am a bit over-optimistic when I say 3.8 27B 's non-thinking output is that good. I spotted a couple of subtle errors in my tests that Low thinking didn't fail on. It is good, but it is not quite Qwen 3.6 35B thinking level. | ||