CoCa is an AI model developed by Google Research (United States), first published in June 2022. It works in the vision domain, on tasks such as image classification, visual question answering and image captioning.
Training it took an estimated 7.3×10²² FLOP of compute (estimation method: hardware). The model has 2,100,000,000 parameters. It was trained on roughly 1.4T datapoints. Training ran on 2,048 Google TPU v4 for about 120 hours. The compute alone is estimated at $78K in 2023 dollars.
Access: Unreleased. Its weights are not openly released. The reference paper has 1,719 citations. Epoch AI rates the confidence of this record as confident.