← Back to context Comment by amelius 7 hours ago For training or for inference? 2 comments amelius Reply ricardobeat 7 hours ago They don't publish numbers, but Anthropic has a single DC with 200k+ GPUs for inference, GPT-6 Astra is said to have trained on 100k+ GPUs. anvuong 6 hours ago Both, especially for training. Astra and Fable were presumably trained on cluster of 100,000k GPUs, or at least a couple of 10Ks.3,800 GPUs is nothing in the frontier side.
ricardobeat 7 hours ago They don't publish numbers, but Anthropic has a single DC with 200k+ GPUs for inference, GPT-6 Astra is said to have trained on 100k+ GPUs.
anvuong 6 hours ago Both, especially for training. Astra and Fable were presumably trained on cluster of 100,000k GPUs, or at least a couple of 10Ks.3,800 GPUs is nothing in the frontier side.
They don't publish numbers, but Anthropic has a single DC with 200k+ GPUs for inference, GPT-6 Astra is said to have trained on 100k+ GPUs.
Both, especially for training. Astra and Fable were presumably trained on cluster of 100,000k GPUs, or at least a couple of 10Ks.
3,800 GPUs is nothing in the frontier side.