AINewsnow

Fireworks post-trained Kimi K3 to reason with ~40% fewer tokens. Production A/B: 49.3K to 29.9K output tokens per task, score 0.751 to 0.753

Fireworks AI released Ember-1 , a post-trained version of Kimi K3. The goal is shorter reasoning traces without losing accuracy. The interesting part is the approach. Fireworks says turning down K3's reasoning effort gave up too much quality. So they trained the model to reason more efficiently ins…

Read the full story at r/machinelearningnews ↗

Timeline · 1 report

  1. 2026-09-28 07:36 · r/machinelearningnews
    Fireworks post-trained Kimi K3 to reason with ~40% fewer tokens. Production A/B: 49.3K to 29.9K output tokens per task, score 0.751 to 0.753

More stories

  1. Kimi K3: A Claude clone or something else? — CoreWeave Blog
  2. Get Started with Kimi K3 on CoreWeave Dedicated Inference — CoreWeave Blog
  3. China's Kimi AI models bypass guardrails on bioweapons, assassinations: 5 key things to know — Mint AI
  4. Chinese AI tool told researchers how to make bioweapons — BBC Technology
  5. Kimi K3.1 identifier reportedly surfaces in Moonshot API registry — TechNode
  6. Claude Opus helped me implement what I have long been trying — r/ClaudeAI
  7. Putting One Kimi Among Four Claudes, Can the Claudes Identify Kimi? — r/ClaudeAI
  8. Why are Gemini's answers outdated? — r/Bard

Get the daily brief of stories like this at 6:30 every morning →