Action Forcing: Training World Models on Unsupervised Video by Recovering Underlying Egomotion Bases
arXiv:2609.30595v1 Announce Type: new Abstract: Synchronised action annotations are needed to train controllable world models and these datasets remain elusive. Existing approaches make use of instrumented platforms with calibrated sensors, costly manual annotation, or latent-action models which la…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-09-28 04:00 · arXiv cs.CV
Action Forcing: Training World Models on Unsupervised Video by Recovering Underlying Egomotion Bases