| ▲ | dzonga 3 days ago | |||||||
to me the biggest event more than deepseek launch was when zAI served their latest model on all Chinese chips. | ||||||||
| ▲ | genxy 3 days ago | parent [-] | |||||||
Model serving is trivial, and inference is just memory bandwidth. The cost of serving will be asymptotic to flash read energy. Having trained on your own chips, that is the impressive part. | ||||||||
| ||||||||