| ▲ | kennywinker an hour ago | |
https://artificialanalysis.ai/?models=qwen3-8-27b%2Cclaude-o... Index methodologies here: https://artificialanalysis.ai/evaluations/artificial-analysi... Also see some specific benchmarks here: https://huggingface.co/Qwen/Qwen3.8-27B e.g. qwen scores 61.7 on swe bench pro, while opus 4.6 scores 53.4. If you want to argue with the benchmarks, go for it. Fwiw i am not saying qwen3.8-27b is better or as good as the frontier. But i am saying it has crossed the threshold and is now a useful tool for coding and debugging. From my experience, Qwen3.6-35b-a3b was what you describe - it could do surgical edits only. | ||