Live
AI models

Gemini 2.5 Flash Native Audio

Published
Jun 3, 2025

Gemini 2.5 Flash Native Audio is an AI model developed by Google DeepMind (United States), first published in June 2025. It works in the speech domain, on tasks such as speech-to-speech, audio question answering, text-to-speech (tts) and speech synthesis.

Epoch AI has no training-compute estimate for this model.

Access: API access. Its weights are not openly released.

Full record
Organization
Google DeepMind
Country of organization
United States
Domain
Speech
Task
Speech-to-speech, Audio question answering, Text-to-speech (TTS), Speech synthesis
Model accessibility
API access
Open weights
No
More from Google DeepMind
SourceEpoch AI, 'AI Models'. Published online at epoch.ai. Retrieved 2026-07-29 from https://epoch.ai/data/ai-models. Licensed under CC BY 4.0.
← All ai models