AINewsnow

Microsoft trained a 4B coding agent almost entirely with Reinforcement Learning, without a bigger teacher

This story is from 2026-09-10. It is preserved in the archive; the latest stories are on the live feed.

A new report called FrogNano makes a claim worth understanding. It comes from the Froggy Team at Microsoft Research Montréal, working with collaborators from Mila and UC San Diego (Kim, Shi et al., arXiv:2609.07925 [cs.AI]). The model is small: 4 billion parameters, far smaller than today's leading…

Read the full story at r/ArtificialInteligence ↗

Timeline · 2 reports

  1. 2026-09-10 14:11 · r/reinforcementlearning
    Microsoft trained a 4B coding agent almost entirely with Reinforcement Learning, without a bigger teacher
  2. 2026-09-10 13:46 · r/ArtificialInteligence
    Microsoft trained a 4B coding agent almost entirely with Reinforcement Learning, without a bigger teacher

More stories

  1. Microsoft exec called AI scraping the “largest theft of labor in human history” — Ars Technica AI
  2. OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web — The Verge AI
  3. Implementing defense-in-depth authorization for MCP tools on Amazon Quick — AWS Machine Learning Blog
  4. Migrating the GitHub Copilot runtime to Rust, using Copilot — GitHub Blog
  5. OpenAI's latest AI revelation is a 'serious situation,' Microsoft's Suleyman tells CNBC — CNBC Technology
  6. A zero-click RCE flaw in AI coding agents could have exposed enterprise systems — InfoWorld AI
  7. Uncontrolled AI could lead to 'silicon species' rivalling humans, warns Microsoft — BBC Technology
  8. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →