| ▲ | frwickst 9 hours ago | |||||||||||||||||||||||||||||||
I'm getting 6.55t/s using the Qwen3.5-397B-A17B-4bit model with the command: ./infer --prompt "Explain quantum computing" --tokens 100 MacBook Pro M5 Pro (64GB RAM) | ||||||||||||||||||||||||||||||||
| ▲ | j45 8 hours ago | parent | next [-] | |||||||||||||||||||||||||||||||
Appreciate the data point. M5 Max would also be interesting to see once available in desktop form. | ||||||||||||||||||||||||||||||||
| ▲ | logicallee 9 hours ago | parent | prev [-] | |||||||||||||||||||||||||||||||
can you post the final result (or as far as you got before you killed it) to show us how cohesive and good it is? I'd like to see an example of the output of this. | ||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||