Google has launched Gemini 3.5 Transcribe, an advanced speech-to-text model designed to improve voice input accuracy and speed. The model supports over 85 languages, automatically corrects verbal stumbles like "ums" and self-corrections, and offers significantly lower latency compared to its predecessor, Chirp 3. Gemini 3.5 Transcribe is available via API endpoints for both real-time streaming and pre-recorded audio processing, with different feature sets and pricing for each. AI
IMPACT This release enhances AI capabilities in voice processing, potentially improving user interfaces and accessibility across various applications.
RANK_REASON Google AI model release with system card.
- Gemini 3.5 Transcribe
- Chirp 3
- Gboard
- Gemini 3.5 Pro
- Pixel 11
- Gemini API
- Gemini Enterprise Agent Platform
AI-generated summary · Google Gemini · from 13 sources. How we write summaries →