Qwen2-Audio is an AI model developed by Alibaba (China), first published in July 2024. It works in the multimodal, language and audio domain, on tasks such as speech recognition (asr), speech synthesis, translation and speech-to-speech.
Epoch AI has no training-compute estimate for this model. The model has 8,200,000,000 parameters.
Access: Open weights (unrestricted). Its weights are openly available. It is built on top of Qwen-7B. Epoch AI rates the confidence of this record as likely.