Google has released Gemini 3.5 Transcribe, an AI speech-to-text model that automatically filters out filler words and corrects speech stumbles in real time. The model achieves a 2.6% word error rate on pre-recorded audio and a 4.0% rate on live streams, delivering transcriptions 70% faster than its predecessor, Chirp 3. It supports over 85 languages and is available via two distinct developer endpoints, powering consumer features like Gboard's Rambler on the Pixel 11 and the macOS Gemini app.
Sep 1, 2026 · 1 source
Sep 1, 2026 · 3 sources
Aug 31, 2026 · 2 sources
Aug 27, 2026 · 5 sources
Story comments
Loading comments…