AINewsnow

Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages

This story is from 2026-08-28. It is preserved in the archive; the latest stories are on the live feed.

Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages They shipped two models instead of one. gemini-3.5-transcribe-live gives sub-second streaming over WebSockets but no diarization and no word-level timestamps, capped at 10-minute session…

Read the full story at r/machinelearningnews ↗

Timeline · 2 reports

  1. 2026-08-29 14:30 · MarkTechPost
    Google AI Releases Gemini Omni 1.1 Flash: 40-Second Scene Extension, First/Last Frame Control, and 4K Upscaling
  2. 2026-08-28 05:39 · r/machinelearningnews
    Google AI Releases Gemini 3.5 Transcribe: A Speech-to-Text Model Reporting 2.6% Average WER Across 85+ Languages

More stories

  1. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  2. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  3. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  4. AI skills — r/AI_Agents
  5. Gemini 4 Pro vs Fable 5 vs GPT6 Astra — r/GeminiAI
  6. Is Gemini 4.0 Pro genuinely coming in October or is it a 3.9 Flash release? Or a surprise 3.5 Pro release? — r/GeminiAI
  7. Plugin4Shell and NIST IR 8587, days apart: what actually authorizes an AI agent’s action? — r/AI_Agents
  8. Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI

Get the daily brief of stories like this at 6:30 every morning →