Google Releases Gemini 3.5 Transcribe Speech-to-Text Model
This story is from 2026-08-26. It is preserved in the archive; the latest stories are on the live feed.
Google launched Gemini 3.5 Transcribe, a speech-to-text model with 2.6% average word error rate across 85+ languages, offered as two endpoints: a streaming version with sub-second transcription and a non-streaming version with speaker diarization.
Read the full story at Ars Technica AI ↗
Timeline · 7 reports
- 2026-08-29 14:30 · MarkTechPost
Google AI Releases Gemini Omni 1.1 Flash: 40-Second Scene Extension, First/Last Frame Control, and 4K Upscaling - 2026-08-28 05:39 · r/machinelearningnews
Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages - 2026-08-28 05:00 · MarkTechPost
Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages - 2026-08-27 20:01 · r/Bard
I tried Gemini 3.5 Transcribe - 2026-08-26 20:21 · r/singularity
Introducing Gemini 3.5 Transcribe - 2026-08-26 19:19 · Ars Technica AI
Google announces Gemini 3.5 Transcribe for AI-powered speech-to-text - 2026-08-26 17:15 · r/GeminiAI
Google just released its new speech-to-text model, Gemini 3.5 Transcribe.