| ▲ | Show HN: Qwen3.6-35B-A3B on a 16 GB M1 Pro with SSD-streamed MoE(github.com) | |||||||
| 22 points by andreaborio 4 days ago | 7 comments | ||||||||
| ▲ | tmzt 3 days ago | parent | next [-] | |||||||
This looks useful for somebody with a 16-32GB Mac Mini interested in running larger MoE models. I've been working on a mesh environment that relies on an explicit prefix hash in the request and enables constructing a new session with a cached pre-filled system prompt specific KV-cache beyond what OpenAI-compatible APIs offer. Can you see a feature like that being supported? | ||||||||
| ||||||||
| ▲ | anentropic a day ago | parent | prev | next [-] | |||||||
does it need M1 Pro specifically, or any M1 could work? | ||||||||
| ||||||||
| ▲ | Capitanai 3 days ago | parent | prev | next [-] | |||||||
[flagged] | ||||||||
| ▲ | 3 days ago | parent | prev | next [-] | |||||||
| [deleted] | ||||||||
| ▲ | andreaborio 4 days ago | parent | prev [-] | |||||||
[flagged] | ||||||||