apple/LensVLM-9B · Hugging Face
https://huggingface.co/bartowski/LensVLM-9B-GGUF LensVLM-9B LensVLM is a 9B Vision Language Model (VLM) that scans compressed images of text, then selectively expands only the relevant pages to their uncompressed form via learned tools. Paper: LensVLM: Selective Context Expansion for Compressed Vis…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-23 18:04 · r/LocalLLaMA
apple/LensVLM-9B · Hugging Face