The hidden advantage of mask resampling: a theory of masked autoencoders
arXiv:2610.01578v1 Announce Type: new Abstract: Why can masked prediction learn useful representations that unmasked reconstruction misses? We study this question in a high-dimensional model of a masked autoencoder (MAE) trained on data with shared latent structure and heterogeneous noise. We prove…
Read the full story at arXiv stat.ML ↗
Timeline · 1 report
- 2026-10-02 04:00 · arXiv stat.ML
The hidden advantage of mask resampling: a theory of masked autoencoders