| ▲ | wongarsu 2 hours ago | ||||||||||||||||||||||
Is frontier scale larger than this? Kimi K3 seems to benchmark in the same range as Opus and Fable. I would have expected they are all in the 2-4T range, with quality of the training and architecture differences as the major differentiators | |||||||||||||||||||||||
| ▲ | porridgeraisin an hour ago | parent [-] | ||||||||||||||||||||||
The number of active parameters is vastly different. Deepseek CEO hinted that he estimates it as an order of magnitude difference in one of his recent interviews. > Seems to benchmark yes, but in human usage the differences show up | |||||||||||||||||||||||
| |||||||||||||||||||||||