| ▲ | andy99 a day ago | ||||||||||||||||
Qwen 3.5 to 3.6 was a big jump for the same size, e.g. 29 to 32 on artificial analysis intelligence for the 35BA3B models. Although I don’t think anyone has released a better model of that size since. I would love to see something like a 90B A6B model that is optimized for 128GB machines e.g. strix halo, I haven’t seen anything really targeting the combination of RAM and compute these machines have, but I’m biased because I have one. | |||||||||||||||||
| ▲ | pixelpoet a day ago | parent [-] | ||||||||||||||||
Yes, yes, yes! I'm absolutely ready and waiting with dual Strix Halo machines here and really want something approaching Opus at home. Speed is secondary concern for now, that would absolutely change the world. Qwen 3.6 27b 8b quant 16b kv cache is already pretty good on the Strix. | |||||||||||||||||
| |||||||||||||||||