Remix.run Logo
QuantumNoodle 4 hours ago

The thing is, AI is a productivity boost but staff drive all the productivity. The catch 22 is AI is expensive so in order to provide it to everyone they need to reduce staffing. The result is remaining folks are able to produce more but the organization as a whole is not exceeding pre-layoff output. And now things are being dropped on the floor and falling in between the cracks -- which impedes efficiency and velocity. Happening at my company now.

AI inference needs to get cheaper or there will always be this "terminal velocity." I suspect AI labs' incentives to reduce costs is only where there is overlap to free up hardware/utilization (to then provide inference to more paying customers.) I highly doubt they will want to make things cheaper for users -- they have debts to pay.

Open weight models are the way and forward, it is the only way an organization can truly control costs by self hosting or buying cheaper inference. Relying on closed weight models is a business risk. Kimi models are only marginally worse than opus but significantly better than the bleeding edge of yesterday's sonnet.

4 hours ago | parent [-]
[deleted]