AudioGen is an AI model developed by Meta AI and Hebrew University of Jerusalem (United States and Israel), first published in March 2023. It works in the audio domain, on tasks such as audio generation.
Training it took an estimated 9.5×10²¹ FLOP of compute (estimation method: hardware). The model has 1,000,000,000 parameters. It was trained on roughly 230.4B datapoints. Training ran on NVIDIA A100 for about 168 hours. The compute alone is estimated at $9K in 2023 dollars.
Access: Open weights (non-commercial). Its weights are openly available. The reference paper has 436 citations. Epoch AI rates the confidence of this record as likely.