| ▲ | l72 an hour ago | |
> In particular, Qwen3.8-Max is the official version based on Qwen3.8-2.4T-A95B with more features, such as vision input & non-thinking support, 1M context length by default, official built-in tools, etc. That is unfortunate, that the open weight model doesn't have vision support or the 1M context length... | ||
| ▲ | wren6991 an hour ago | parent [-] | |
People have had surprising success adding vision to open-weight LLMs that ship without it, like DSV4 Flash [1] or GLM-5.2 [2]. Given this model is already vision-trained I expect that approach will work well here. [1] https://old.reddit.com/r/LocalLLaMA/comments/1vl6ior/i_gave_... | ||