Evolving a Mario controller with NEAT: three failed replays and a staircase that finally gets beaten
This story is from 2026-09-16. It is preserved in the archive; the latest stories are on the live feed.
This gap near the end of Bowser’s Crown 4-1 kept catching different controllers. The clip shows three selected earlier genome replays, followed by the controller that clears the level. These aren’t consecutive training attempts, and none of the controllers is learning during its replay. I’m continu…
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-09-16 13:25 · r/reinforcementlearning
Evolving a Mario controller with NEAT: three failed replays and a staircase that finally gets beaten