GPT-4 (Mar 2023) (OpenAI) and Minerva (540B) (Google) are both frontier AI models. GPT-4 (Mar 2023) was published in March 2023 and Minerva (540B) in June 2022.
GPT-4 (Mar 2023) was trained on 2.1 x 10^25 FLOP, about 7.7x the compute of Minerva (540B) at 2.7 x 10^24 FLOP. Training compute is the closest available proxy for how much was invested in a model, though it says nothing on its own about how well that compute was spent.
These two models share no benchmark on which both have been scored, so no direct performance comparison is possible here. The specification table below is a comparison of inputs, not of results.
| Field | GPT-4 (Mar 2023) | Minerva (540B) |
|---|---|---|
| Organization | OpenAI | |
| Published | Mar 15, 2023 | Jun 29, 2022 |
| Training compute | 2.1 x 10^25 FLOP | 2.7 x 10^24 FLOP |
| Parameters | 1.8T | 540.4B |
| Dataset size | 5.4T | 26B |
| Training hardware | NVIDIA A100 SXM4 40 GB | Google TPU v4 |
| Chips used | 25,000 | 1,024 |
| Training time | 2.3K h | 696 h |
| Training cost (2023 USD) | $37M | — |
| Training power draw | 19.9 MW | 698.4 kW |
| Accessibility | API access | Unreleased |
| Open weights | No | No |
| Country | United States | United States |