| ▲ | shangofox an hour ago | |
I mean even if it take a few months, it'll still be out of date. But there was a hypothetical when it came up in Feb, would you want Qwen 3.5 at like 10k tokens per second. At the time people were no doubt saying yes but now 3.8 is out, is that still desirable? | ||
| ▲ | xienze an hour ago | parent [-] | |
There's soooo much stuff that such a model is still capable of doing in the pursuit of getting a better overall answer. Imagine a powerful research agent that blasts out dozens of the small, cheap models to fetch and summarize one page each. Then the beefy researcher model performs the final analysis. | ||