| ▲ | fooker a day ago | |||||||||||||||||||||||||||||||
Economies of scale. You need a cluster of 8-12 H100s to run the largest models locally. It doesn't make sense to run these locally yet unless your use case also involves making it available for several dozen concurrent users. | ||||||||||||||||||||||||||||||||
| ▲ | kingleopold a day ago | parent [-] | |||||||||||||||||||||||||||||||
not to miss, future models will be more compute hungry too. Current hardware prices are still goin up and no it's not cheaper to run your AI for like %99 of the people because of lots of costs, it's not just hardware. | ||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||