| ▲ | kennywinker an hour ago | |||||||
It literally is usable now. A 5060 for $800 can run qwen3.8-27b 4bit at >40t/s, and the model beats opus 4.6 (max). | ||||||||
| ▲ | TomBombadildoze an hour ago | parent [-] | |||||||
Beats Opus 4.6 at what exactly? It certainly isn't code. I use a combination of a Claude Max subscription and local inference, including qwen3.8-27b, 4bit. I have found qwen to be absolutely useless at anything but very specific, surgical code changes. In my experience, for anything even remotely nuanced, a frontier model is required. | ||||||||
| ||||||||