GRAFT: Growing Agglomerative Foundation Models via Continual Teacher Distillation
arXiv:2610.02597v1 Announce Type: new Abstract: Vision foundation models such as DINOv2, SigLIP2, and MASt3R develop complementary capabilities from different pretraining objectives, yet their knowledge remains distributed across separate, specialized models. Multi-teacher knowledge distillation of…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-10-05 04:00 · arXiv cs.CV
GRAFT: Growing Agglomerative Foundation Models via Continual Teacher Distillation