introducing Gemini 3.5 Transcribe: our most precise speech-to-text model yet, designed for intelligent voice interactions
this new model:
- seamlessly handles self-corrections
- removes filler words to deliver clean, formatted text
- understands your natural intent and speaking style
- captures audio, even in noisy environments
- supports 85+ languages, including regional accents and diverse dialects
try it today in the Gemini API and in AI Studio