Hacker News
new
past
comments
ask
show
jobs
points
by
eigenspace
7 hours ago
|
comments
by
amelius
6 hours ago
|
[-]
For training or for inference?
reply
by
ricardobeat
5 hours ago
|
parent
|
next
[-]
They don't publish numbers, but Anthropic has a single DC with 200k+ GPUs for inference, GPT-6 Astra is said to have trained on 100k+ GPUs.
reply
by
anvuong
5 hours ago
|
parent
|
prev
|
[-]
Both, especially for training. Astra and Fable were presumably trained on cluster of 100,000k GPUs, or at least a couple of 10Ks.
3,800 GPUs is nothing in the frontier side.
reply