实时
Head to head

Grok 3 vs Llama 3.1-405B

xAI
Grok 3
February 2025
vs
Meta AI
Llama 3.1-405B
July 2024
3.5×10²⁶Training compute (FLOP)3.8×10²⁵
139Capability index129
$218MTraining cost$53M
Benchmarks won (1 scored on both)
1 Grok 3Llama 3.1-405B 0

Grok 3 (xAI) and Llama 3.1-405B (Meta AI) are both frontier AI models. Grok 3 was published in February 2025 and Llama 3.1-405B in July 2024.

Grok 3 was trained on 3.5×10²⁶ FLOP, about 9.2x the compute of Llama 3.1-405B at 3.8×10²⁵ FLOP. Training compute is the closest available proxy for how much was invested in a model, though it says nothing on its own about how well that compute was spent.

Grok 3 wins the single benchmark both were scored on. Benchmark counts are a crude scoreboard — the evaluations differ wildly in what they measure and in how saturated they are — so the per-benchmark table below matters more than the tally.

Specifications
xAI
Organization
Meta AI
Feb 17, 2025
Published
Jul 23, 2024
3.5×10²⁶ FLOP
Training compute9.2x
3.8×10²⁵ FLOP
3T
Parameters7.4x
405B
--
Dataset size
15.6T
NVIDIA H100 SXM5 80GB
Training hardware
NVIDIA H100 SXM5 80GB
80,000
Chips used4.9x
16,384
2,160 h
Training time
2,142 h
$218M
Training cost (2023 USD)4.1x
$53M
109.9 MW
Training power draw
22.6 MW
API access
Accessibility
Open weights (restricted use)
No
Open weights
Yes
United States
Country
United States
Shared benchmarks
Related comparisons
SourceEpoch AI, 'AI Models'. Published online at epoch.ai. Retrieved 2026-07-29 from https://epoch.ai/data/ai-models. Licensed under CC BY 4.0.