DD-PPO is an AI model developed by Georgia Institute of Technology, Facebook AI Research, Oregon State University and Simon Fraser University (United States, France and Canada), first published in December 2019. It works in the robotics domain, on tasks such as object detection.
Training it took an estimated 7.8×10²⁰ FLOP of compute (estimation method: hardware). It was trained on roughly 2.5B datapoints. Training ran on 64 NVIDIA V100 for about 66 hours. The compute alone is estimated at $2K in 2023 dollars.
Access: Unreleased. Its weights are not openly released. The reference paper has 611 citations. Epoch AI rates the confidence of this record as likely.