Comment by eigenspace
8 hours ago
That cluster is literally orders of magnitude smaller than the compute pools used by Anthropic or OpenAI.
8 hours ago
That cluster is literally orders of magnitude smaller than the compute pools used by Anthropic or OpenAI.
For training or for inference?
They don't publish numbers, but Anthropic has a single DC with 200k+ GPUs for inference, GPT-6 Astra is said to have trained on 100k+ GPUs.
Both, especially for training. Astra and Fable were presumably trained on cluster of 100,000k GPUs, or at least a couple of 10Ks.
3,800 GPUs is nothing in the frontier side.