AINewsnow

Google's RRSI lets an agent rewrite its own harness with frozen weights: Terminal-Bench 2.1 74.2% → 80.2%, 6/6 held-out benchmarks up, Apache 2.0

Google Research released Regularized Recursive Self-Improvement (RRSI): an LLM agent that edits its own harness (prompts, tools, memory, control flow, sub-agents) while the model weights stay frozen. The regularization is what stops the usual failure mode, where a self-improving loop overfits the t…

Read the full story at r/machinelearningnews ↗

Timeline · 1 report

  1. 2026-09-29 09:19 · r/machinelearningnews
    Google's RRSI lets an agent rewrite its own harness with frozen weights: Terminal-Bench 2.1 74.2% → 80.2%, 6/6 held-out benchmarks up, Apache 2.0

More stories

  1. NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring — NVIDIA Technical Blog
  2. Opus 5.5 — r/ClaudeAI
  3. Gemini 4.0 where you at — r/GeminiAI
  4. See what 4 builders are making with Gemini 3.8 Flash — Google Gemini Blog
  5. 3 ways this grocer cooks for 200 guests with Gemini — Google Gemini Blog
  6. Bill Gates warns of AI risks, calls for Congress to step in: 'It's not a hoax at all' — Mint AI
  7. How does Google feel when they no longer see their name in benchmarks? really sad — r/GeminiAI
  8. Optimizing my AI subscriptions: Claude Pro (Opus) vs. ChatGPT Plus vs. Perplexity Pro? — r/AI_Agents

Get the daily brief of stories like this at 6:30 every morning →