Dreamer needed 5M env steps where D4PG needed 1e8 and A3C 1e9. I wrote up how "learning in imagination" went from a 2018 paper to a billion-dollar bet.
This story is from 2026-09-13. It is preserved in the archive; the latest stories are on the live feed.
Not a paper, a long-form explainer I wrote for people outside RL, but it's built on the primary sources this sub will know: Ha & Schmidhuber's World Models, Hafner's Dreamer / DreamerV3, DayDreamer on the real quadruped (walking in ~1 hour, no sim), through to Genie 3 and JEPA. The bit I'd genuinel…
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-09-13 00:20 · r/reinforcementlearning
Dreamer needed 5M env steps where D4PG needed 1e8 and A3C 1e9. I wrote up how "learning in imagination" went from a 2018 paper to a billion-dollar bet.