Fireworks post-trained Kimi K3 to reason with ~40% fewer tokens. Production A/B: 49.3K to 29.9K output tokens per task, score 0.751 to 0.753
Fireworks AI released Ember-1 , a post-trained version of Kimi K3. The goal is shorter reasoning traces without losing accuracy. The interesting part is the approach. Fireworks says turning down K3's reasoning effort gave up too much quality. So they trained the model to reason more efficiently ins…
Read the full story at r/machinelearningnews ↗
Timeline · 1 report
- 2026-09-28 07:36 · r/machinelearningnews
Fireworks post-trained Kimi K3 to reason with ~40% fewer tokens. Production A/B: 49.3K to 29.9K output tokens per task, score 0.751 to 0.753