HN
Hacker News
Top
New
Best
Ask
Show
Jobs
Comment by amelius | Hacker News Reader
Parent
Full thread
amelius
·
For training or for inference?
View on HN
ricardobeat
·
They don't publish numbers, but Anthropic has a single DC with 200k+ GPUs for inference, GPT-6 Astra is said to have trained on 100k+ GPUs.
anvuong
·
Both, especially for training. Astra and Fable were presumably trained on cluster of 100,000k GPUs, or at least a couple of 10Ks.
3,800 GPUs is nothing in the frontier side.
Reply on news.ycombinator.com