| ▲ | ignoramous 4 hours ago | |||||||
Per Xiaomi, MiMo v2.6 training run cost $3.47m. A far cry from the estimated costs ($100m+) for the Big 5 (MSL, xAI, GDM, OAI, Ant). I wouldn't be surprised if salaries and R&D costs have similar drastic disparities. For a model that matches Muse Spark 1.3 in benchmarks, MiMo v2.6 Pro is incredibly cheap, given its cache rates will remain $0.0036 per million. | ||||||||
| ▲ | imjonse 3 hours ago | parent | next [-] | |||||||
That is the RL training cost only. Their announcement blog mentions this: https://mimo.xiaomi.com/mimo-v2-6#scaling-rl-fully-open-sour... | ||||||||
| ▲ | drbscl 3 hours ago | parent | prev | next [-] | |||||||
My understanding of tech salaries in China is that they are pretty decent, but not as high as in SF; closer to typical European salaries. Mostly due to lower cost of living; Shenzhen is way cheaper than SV | ||||||||
| ||||||||
| ▲ | NortySpock 4 hours ago | parent | prev [-] | |||||||
I sorta got the impression that the $3.47 million only covered post-training , given that few of the graphs start at zero. Is a barely-trained model going to score 48 on DeepSWE v1.1 ? | ||||||||