Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages
This story is from 2026-08-28. It is preserved in the archive; the latest stories are on the live feed.
Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages They shipped two models instead of one. gemini-3.5-transcribe-live gives sub-second streaming over WebSockets but no diarization and no word-level timestamps, capped at 10-minute session…
Read the full story at r/machinelearningnews ↗
Timeline · 2 reports
- 2026-08-29 14:30 · MarkTechPost
Google AI Releases Gemini Omni 1.1 Flash: 40-Second Scene Extension, First/Last Frame Control, and 4K Upscaling - 2026-08-28 05:39 · r/machinelearningnews
Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages