Live
AI models

OpenOmni

Parameters
7B
Published
May 24, 2025

OpenOmni is an AI model developed by Chinese Academy of Sciences, Shenzhen Institute of Advanced Technology, University of Chinese Academy of Sciences, National University of Singapore and University of Science and Technology of China (USTC) (China and Singapore), first published in May 2025. It works in the multimodal, language, vision and speech domain, on tasks such as speech-to-text, speech recognition (asr), image captioning and 5 more.

Epoch AI has no training-compute estimate for this model. The model has 7,000,000,000 parameters. Training ran on 8 NVIDIA A100.

Access: Open weights (unrestricted). Its weights are openly available. It is built on top of Qwen2.5 Instruct (7B),CLIP (ViT L/14@336px),Whisper v3. Epoch AI rates the confidence of this record as confident.

Full record
Organization
Chinese Academy of Sciences, Shenzhen Institute of Advanced Technology, University of Chinese Academy of Sciences, National University of Singapore, University of Science and Technology of China (USTC)
Country of organization
China, Singapore
Domain
Multimodal, Language, Vision, Speech
Task
Speech-to-text, Speech recognition (ASR), Image captioning, Visual question answering, Language modeling/generation, Question answering, Text-to-speech (TTS), Speech synthesis
Compute estimation method
Hardware
Parameters
7,000,000,000
Training hardware
NVIDIA A100
Chips used
8
Chip-hours
664
Training power draw
6.3 kW
Model accessibility
Open weights (unrestricted)
Open weights
Yes
Base model
Qwen2.5 Instruct (7B), CLIP (ViT L/14@336px), Whisper v3
Epoch confidence
Confident
SourceEpoch AI, 'AI Models'. Published online at epoch.ai. Retrieved 2026-07-29 from https://epoch.ai/data/ai-models. Licensed under CC BY 4.0.
← All ai models