In Q3 2024, Meta trained a 70B parameter code-specialized LLM on 100,000 Nvidia H100 GPUs, achieving 214 TFLOPS per GPU and 92% cluster utilization – a 3x improvement over their 2023 16k A100 cluster runs, with total training cost of $17.4M for 21 days of continuous operation.
📡 Hacker News Top Stories Right Now
Ghostty is leaving...
🛡️ VERIFIED CYBER INTELLIGENCE ID: #3445429