CamemBERT is an AI model developed by Facebook, INRIA and Sorbonne University (United States and France), first published in November 2019. It works in the language domain, on tasks such as language modeling/generation, part-of-speech tagging and named entity recognition (ner).
Training it took an estimated 8.3×10²⁰ FLOP of compute (estimation method: hardware,operation counting). The model has 335,000,000 parameters. It was trained on roughly 28.6B datapoints. Training ran on NVIDIA V100 for about 24 hours. The compute alone is estimated at $2K in 2023 dollars.
Access: Open weights (unrestricted). Its weights are openly available. The reference paper has 1,083 citations. Epoch AI rates the confidence of this record as confident.