World model for better results on ur current stack
Hey guyss, Ive been working on this for few months, and finally getting some good results, so figured I'd share. A coding agent only sees the run it's currently in. It doesn't really know how runs like this usually go, or what kind of repo it's working inside. So I trained a small model on 50K+ rea…
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-09-20 09:33 · r/reinforcementlearning
World model for better results on ur current stack