GPT-3.5 (davinci-002) (OpenAI) and Wu Dao 2.0 (Beijing Academy of Artificial Intelligence / BAAI) are both frontier AI models. GPT-3.5 (davinci-002) was published in March 2022 and Wu Dao 2.0 in May 2021.
GPT-3.5 (davinci-002) was trained on 2.6 x 10^24 FLOP, about 1.7x the compute of Wu Dao 2.0 at 1.5 x 10^24 FLOP. Training compute is the closest available proxy for how much was invested in a model, though it says nothing on its own about how well that compute was spent.
These two models share no benchmark on which both have been scored, so no direct performance comparison is possible here. The specification table below is a comparison of inputs, not of results.
| Field | GPT-3.5 (davinci-002) | Wu Dao 2.0 |
|---|---|---|
| Organization | OpenAI | Beijing Academy of Artificial Intelligence / BAAI |
| Published | Mar 15, 2022 | May 31, 2021 |
| Training compute | 2.6 x 10^24 FLOP | 1.5 x 10^24 FLOP |
| Parameters | — | 1.8T |
| Dataset size | — | 4.9T |
| Training hardware | NVIDIA A100 SXM4 40 GB | — |
| Training cost (2023 USD) | $5M | $3M |
| Accessibility | API access | API access |
| Open weights | No | No |
| Country | United States | China |