| ▲ | mmastrac 3 hours ago | |
The comparisons with other models here are odd.. the other models change depending on the task. It would be far more useful to at least compare against the more recent open models (DS4Flash/GLM53Flash/Qwen38). | ||
| ▲ | cogman10 an hour ago | parent [-] | |
They are trying to keep the models within the same quant class, which is tough to do since a lot of models aren't distilled to lower quants. There is, for example, no Qwen3.8 7B. It is odd to me, though, that they didn't run the same benchmark suite for the various quants. | ||