Remix.run Logo
mindwok 2 hours ago

Does anyone else feel like the writing is on the wall for a future of local models? Spamming data centres everywhere, powering them, having to commit insane capital to hardware, all the effort to serve inference over a network reliably - when here we are with a frontier model nearly running on a laptop.

Local AI on your device seems like a much more likely future to me than datacenters in space. For inference at least, training is another story.

nomel 2 hours ago | parent [-]

0.01 tk/s on an M1 Max is not "nearly". This is completely unusable, and in no way cost effective.