Reddit r/LocalLLaMASeptember 23, 2026
apple/LensVLM-9B · Hugging Face
Excerpt
https://huggingface.co/bartowski/LensVLM-9B-GGUF LensVLM-9B LensVLM is a 9B Vision Language Model (VLM) that scans compressed images of text, then selectively expands only the relevant pages to their uncompressed form via learned tools. Paper: LensVLM: Selective Context Expansion for Compressed Visual Representation of Text Code: https://github.com/apple-aiml-research/ml-lensvlm License All ML model files in this repository, including Apple's modifications to the Qwen model, are provided under t