My agent is a stuntman.
This story is from 2026-09-08. It is preserved in the archive; the latest stories are on the live feed.
By the way do anyone know any algorithm or setup that can achieve over 500 on average on HopperHop in DCM in less then 500K steps or 1M steps?
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-09-08 13:18 · r/reinforcementlearning
My agent is a stuntman.