Google releases Gemini 3.5 Transcribe in public preview

The speech-to-text model is available in Google AI Studio, the Gemini API, and Gemini Enterprise Agent Platform for streaming and pre-recorded audio across 85+ languages.

Developers building voice agents and transcription pipelines gain higher-accuracy speech recognition that natively cleans disfluencies without auxiliary formatting pipelines.

Google releases Gemini 3.5 Transcribe in public preview

Sources

Read this as text

Back to the AI news