Step-Audio-TTS-3B is an AI model developed by StepFun (China), first published in February 2025. It works in the speech domain, on tasks such as text-to-speech (tts) and speech synthesis.
Epoch AI has no training-compute estimate for this model. The model has 3,000,000,000 parameters.
Access: Open weights (unrestricted). Its weights are openly available.