Fish-Speech 1.4 is an AI model developed by Fish Audio (United States), first published in November 2024. It works in the speech domain, on tasks such as speech synthesis and text-to-speech (tts).
Training it took an estimated 1.9×10²¹ FLOP of compute (estimation method: hardware). It was trained on roughly 500B datapoints. Training ran on 8 NVIDIA H100 SXM5 80GB,NVIDIA GeForce RTX 4090 for about 168 hours.
Access: Open weights (non-commercial). Its weights are openly available. Epoch AI rates the confidence of this record as confident.