News
Google releases Gemini 3.5 Transcribe
Google has released Gemini 3.5 Transcribe, a speech-to-text model that offers APIs for both real-time streaming and recorded files, and supports more than 85 languages.
Google has released the Gemini 3.5 Transcribe speech-to-text model and made it available for public preview through the Gemini API and Enterprise Agent platform.
The model processes real-time streaming speech using gemini-3.5-transcribe-live and pre-recorded audio using gemini-3.5-transcribe, supporting speaker recognition, word-level timestamps, and more than 85 languages.
refer to:Google's official announcement