AINewsnow

My AI learns to clear Super Mario Bros 1-1 in 15 mins and it is not PPO based

I tested Adapt-1, a non-LLM learning and reasoning system by Rei Labs, by having it learn and play Super Mario Bros, and it performed quite well. I tried it on World 1-1, starting untrained. It learned a reactive policy from its own play in about 36 minutes of gameplay, then cleared the level with…

Read the full story at r/reinforcementlearning ↗

Timeline · 3 reports

  1. 2026-10-03 07:52 · r/learnmachinelearning
    Training AI to play and clear Super Mario Bros is easier than I thought
  2. 2026-10-03 07:49 · r/deeplearning
    My AI learns to clear Super Mario Bros 1-1 in 15 mins and it is not PPO based
  3. 2026-10-03 07:43 · r/reinforcementlearning
    My AI learns to clear Super Mario Bros 1-1 in 15 mins and it is not PPO based

More stories

  1. NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI — NVIDIA Blog
  2. Gemini 4 Argon: our next era of frontier intelligence — Google Gemini Blog
  3. Guided Vision in Gemini Live: built for accessibility — Google Gemini Blog
  4. Google tests its plan for AI data centers in space with Project Suncatcher — Scientific American
  5. Google announces Gemini 4 Argon AI model, but you can't use it yet — Ars Technica AI
  6. OpenAI DevDay 2026 Keynote (FULL) — OpenAI YouTube
  7. The latest AI news we announced in September 2026 — Google Gemini Blog
  8. Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs — Allen Institute for AI (Ai2)

Get the daily brief of stories like this at 6:30 every morning →