AB
AiBoss
News

Google releases Gemini 3.5 Transcribe

Google has released Gemini 3.5 Transcribe, a speech-to-text model that offers APIs for both real-time streaming and recorded files, and supports more than 85 languages.

Google has released the Gemini 3.5 Transcribe speech-to-text model and made it available for public preview through the Gemini API and Enterprise Agent platform.

The model processes real-time streaming speech using gemini-3.5-transcribe-live and pre-recorded audio using gemini-3.5-transcribe, supporting speaker recognition, word-level timestamps, and more than 85 languages.

refer to:Google's official announcement