Intelligent transcription with Gemini 3.5 Transcribe
Summary
Google has released Gemini 3.5 Transcribe, a new speech-to-text model (AI that converts spoken words into written text) that converts raw audio into accurate, formatted text while handling background noise, technical jargon, and natural speech patterns like self-corrections. The model is available through two APIs (interfaces for developers to build with): a real-time streaming API for interactive voice apps and a pre-recorded audio API for meetings and call logs, with support for over 85 languages and multi-speaker identification.
Classification
Affected Vendors
Related Issues
Original source: https://deepmind.google/blog/intelligent-transcription-with-gemini-3-5-transcribe/
First tracked: August 26, 2026 at 02:00 PM
Classified by LLM (prompt v3) · confidence: 92%