Google Releases Gemini 3.5 Transcribe Speech-to-Text Model
This story is from 2026-08-26. It is preserved in the archive; the latest stories are on the live feed.
Google launched Gemini 3.5 Transcribe, a speech-to-text model with 2.6% average word error rate across 85+ languages, offered as two endpoints: a streaming version with sub-second latency but no speaker diarization, and another endpoint.
Read the full story at Ars Technica AI ↗
Timeline · 4 reports
- 2026-08-29 14:30 · MarkTechPost
Google AI Releases Gemini Omni 1.1 Flash: 40-Second Scene Extension, First/Last Frame Control, and 4K Upscaling - 2026-08-28 05:39 · r/machinelearningnews
Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages - 2026-08-28 05:00 · MarkTechPost
Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages - 2026-08-26 19:19 · Ars Technica AI
Google announces Gemini 3.5 Transcribe for AI-powered speech-to-text