AINewsnow

Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning

Kyutai has released Voice of Reason, 2 open-weight speech-to-speech models built on GLM-4-Voice-9B. Supervised fine-tuning and reinforcement learning lift spoken GSM8K accuracy from 27.3% to 77.1%. There is no transcription step and no text LLM in the loop. Both checkpoints are on Hugging Face and…

Read the full story at MarkTechPost ↗

Timeline · 1 report

  1. 2026-09-23 06:33 · MarkTechPost
    Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning

More stories

  1. Jev's calibration was measured. The LLMs won [D] — r/MachineLearning
  2. 299 real user intents tested Jev against production base line. Here is the result. — r/AI_Agents
  3. MiMo-V2.6-Flash on vLLM: fixes for "empty responses" with thinking + tools, and a hidden 2,048-token output cap — r/LocalLLaMA
  4. GLM 5.3 now available in Mistral Vibe Code for Pro, Team and Enterprise — r/ChatGPTCoding
  5. CAISI’s Assessment of Z.ai’s GLM-5.3 Cyber Capabilities — r/ArtificialInteligence
  6. CPU Inference on Dell Poweredge R820 — r/LocalLLM
  7. Qwen Image 2.1 - More detailed tests — r/StableDiffusion
  8. we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI

Get the daily brief of stories like this at 6:30 every morning →