Small yet Assistive: Spatially-Aware Post-Training for Low Vision
arXiv:2609.28757v1 Announce Type: new Abstract: An estimated 1 billion people worldwide live with vision impairment, yet current vision-language models (VLMs) produce descriptions too vague for safe navigation by blind and low-vision (BLV) users. Large VLMs can generate high-quality audio-descripti…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-09-25 04:00 · arXiv cs.CV
Small yet Assistive: Spatially-Aware Post-Training for Low Vision