NatureLM-audio is an AI model developed by Earth Species Project (United States), first published in November 2024. It works in the audio domain, on tasks such as audio classification.
Training it took an estimated 1.4×10²¹ FLOP of compute (estimation method: hardware). The model has 665,000,000 parameters. Training ran on 8 NVIDIA H100 PCIe for about 216 hours.
Access: Open weights (non-commercial). Its weights are openly available. It is built on top of Llama 3.1-8B. Epoch AI rates the confidence of this record as confident.