Remix.run Logo
dotinvictim an hour ago

local llm don't make sense currently consumer compute is not upto mark it may take atleast 7 more years to be usable

kennywinker 24 minutes ago | parent [-]

It literally is usable now. A 5060 for $800 can run qwen3.8-27b 4bit at >40t/s, and the model beats opus 4.6 (max).

TomBombadildoze 4 minutes ago | parent [-]

Beats Opus 4.6 at what exactly? It certainly isn't code.

I use a combination of a Claude Max subscription and local inference, including qwen3.8-27b, 4bit. I have found qwen to be absolutely useless at anything but very specific, surgical code changes. In my experience, for anything even remotely nuanced, a frontier model is required.