Atom is 60M Param (around 133x to 400x smaller).
16ms latency. And locally run.
https://at0m.pienomial.com/
Why go big when you can go small ?
Cause it's not open?
Good point.
To counter, most of the AI is not open. So is none of Microsoft Products. As long as they work, we keep using them.