Llama Nemotron Ultra 253B (NVIDIA) and PaLM 2 (Google) are both frontier AI models. Llama Nemotron Ultra 253B was published in March 2025 and PaLM 2 in May 2023.
Llama Nemotron Ultra 253B was trained on 3.9 x 10^25 FLOP, about 5.3x the compute of PaLM 2 at 7.3 x 10^24 FLOP. Training compute is the closest available proxy for how much was invested in a model, though it says nothing on its own about how well that compute was spent.
These two models share no benchmark on which both have been scored, so no direct performance comparison is possible here. The specification table below is a comparison of inputs, not of results.
| Field | Llama Nemotron Ultra 253B | PaLM 2 |
|---|---|---|
| Organization | NVIDIA | |
| Published | Mar 18, 2025 | May 10, 2023 |
| Training compute | 3.9 x 10^25 FLOP | 7.3 x 10^24 FLOP |
| Parameters | 253B | 340B |
| Dataset size | 603B | 3.6T |
| Training hardware | — | Google TPU v4 |
| Training cost (2023 USD) | — | $5M |
| Accessibility | Open weights (restricted use) | API access |
| Open weights | Yes | No |
| Country | United States | United States |