Structure-Guided Masked Autoencoders for Ultra-High Resolution Scientific Image Understanding
arXiv:2609.30682v1 Announce Type: new Abstract: Self-supervised pre-training with Vision Transformers, including Masked Autoencoders (MAE), is difficult to apply to gigapixel scientific images. Random masking is poorly matched to the structured, multi-scale morphology of scientific data, while unif…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-09-28 04:00 · arXiv cs.CV
Structure-Guided Masked Autoencoders for Ultra-High Resolution Scientific Image Understanding