Remix.run Logo
▲ eigenspace 5 hours ago

That cluster is literally orders of magnitude smaller than the compute pools used by Anthropic or OpenAI.

▲amelius 4 hours ago | parent [-]

For training or for inference?

▲ricardobeat 3 hours ago | parent | next [-]

They don't publish numbers, but Anthropic has a single DC with 200k+ GPUs for inference, GPT-6 Astra is said to have trained on 100k+ GPUs.

▲anvuong 2 hours ago | parent | prev [-]

Both, especially for training. Astra and Fable were presumably trained on cluster of 100,000k GPUs, or at least a couple of 10Ks.

3,800 GPUs is nothing in the frontier side.