My AI learns to clear Super Mario Bros 1-1 in 15 mins and it is not PPO based
I tested Adapt-1, a non-LLM learning and reasoning system by Rei Labs, by having it learn and play Super Mario Bros, and it performed quite well. I tried it on World 1-1, starting untrained. It learned a reactive policy from its own play in about 36 minutes of gameplay, then cleared the level with…
Read the full story at r/reinforcementlearning ↗
Timeline · 3 reports
- 2026-10-03 07:52 · r/learnmachinelearning
Training AI to play and clear Super Mario Bros is easier than I thought - 2026-10-03 07:49 · r/deeplearning
My AI learns to clear Super Mario Bros 1-1 in 15 mins and it is not PPO based - 2026-10-03 07:43 · r/reinforcementlearning
My AI learns to clear Super Mario Bros 1-1 in 15 mins and it is not PPO based