Inflection-2 (Inflection AI) and Nemotron-4 340B (NVIDIA) are both frontier AI models. Inflection-2 was published in November 2023 and Nemotron-4 340B in June 2024.
Nemotron-4 340B was trained on 1.8 x 10^25 FLOP, about 1.8x the compute of Inflection-2 at 10^25 FLOP. Training compute is the closest available proxy for how much was invested in a model, though it says nothing on its own about how well that compute was spent.
These two models share no benchmark on which both have been scored, so no direct performance comparison is possible here. The specification table below is a comparison of inputs, not of results.
| Field | Inflection-2 | Nemotron-4 340B |
|---|---|---|
| Organization | Inflection AI | NVIDIA |
| Published | Nov 22, 2023 | Jun 14, 2024 |
| Training compute | 10^25 FLOP | 1.8 x 10^25 FLOP |
| Parameters | — | 340B |
| Dataset size | — | 9T |
| Training hardware | NVIDIA H100 SXM5 80GB | NVIDIA H100 SXM5 80GB |
| Chips used | 5,000 | 6,144 |
| Training time | — | 2.2K h |
| Training cost (2023 USD) | $13M | $21M |
| Training power draw | 6.9 MW | 8.5 MW |
| Accessibility | Hosted access (no API) | Open weights (unrestricted) |
| Open weights | No | Yes |
| Country | United States | United States |