Claude 3.5 Sonnet (Anthropic) and Grok-2 (xAI) are both frontier AI models. Claude 3.5 Sonnet was published in June 2024 and Grok-2 in August 2024.
Grok-2 was trained on 3 x 10^25 FLOP, essentially the same compute as Claude 3.5 Sonnet at 2.7 x 10^25 FLOP. Training compute is the closest available proxy for how much was invested in a model, though it says nothing on its own about how well that compute was spent.
Grok-2 wins the single benchmark both were scored on. Benchmark counts are a crude scoreboard — the evaluations differ wildly in what they measure and in how saturated they are — so the per-benchmark table below matters more than the tally.
| Field | Claude 3.5 Sonnet | Grok-2 |
|---|---|---|
| Organization | Anthropic | xAI |
| Published | Jun 20, 2024 | Aug 13, 2024 |
| Training compute | 2.7 x 10^25 FLOP | 3 x 10^25 FLOP |
| Training hardware | — | NVIDIA H100 SXM5 80GB |
| Training cost (2023 USD) | $26M | $32M |
| Accessibility | API access | API access |
| Open weights | No | No |
| Country | United States | United States |
| Benchmark | Claude 3.5 Sonnet | Grok-2 |
|---|---|---|
| Epoch Capabilities Index(ECI Score) | 130 | 130.8 |