| ▲ | wren6991 a day ago | |
> Is there any good reason to believe there is a lot of headroom or there is not? It's hard to answer quantitatively, but for example Qwen3.5 -> 3.6 was a significant step in capability, arising from continued post-training of the same models. If we were at the end of low-parameter-count scaling then that would be a surprising datapoint. | ||