| ▲ | GLM-5.3-Flash-GGUF(huggingface.co) | |
| 9 points by walrus01 15 hours ago | 1 comments | ||
| ▲ | walrus01 15 hours ago | parent [-] | |
Viable GGUFs without excessive loss at Q4 and better for people who have either 256GB or 512GB inference systems. We're seeing a real flurry of 'very capable' open weight models release in the last 3-4 days, including Qwen 3.8-Flash-Next which fits on a 256GB system in Q8. | ||