Google's RRSI lets an agent rewrite its own harness with frozen weights: Terminal-Bench 2.1 74.2% → 80.2%, 6/6 held-out benchmarks up, Apache 2.0
Google Research released Regularized Recursive Self-Improvement (RRSI): an LLM agent that edits its own harness (prompts, tools, memory, control flow, sub-agents) while the model weights stay frozen. The regularization is what stops the usual failure mode, where a self-improving loop overfits the t…
Read the full story at r/machinelearningnews ↗
Timeline · 1 report
- 2026-09-29 09:19 · r/machinelearningnews
Google's RRSI lets an agent rewrite its own harness with frozen weights: Terminal-Bench 2.1 74.2% → 80.2%, 6/6 held-out benchmarks up, Apache 2.0