// HACKER NEWS — CYBERSECURITY
LensVLM: Compressing long context as images, expanding only relevant pages
How to use apple/LensVLM-9B with Docker Model Runner:
LensVLM is a 9B Vision Language Model (VLM) that scans compressed images of text,
then selectively expands only the relevant pages to their uncompressed form via
learned tools.
All ML model files in this repository, including Apple's modifications to the Qwen
model, are provided under the terms of the
Apple Machine Learning Research Model License.
The source code that accompanies this model is distributed separately and is provided
under the terms of the Apple Sample Code License.
Compression options: 5x, 10x, 15x. See the
repository README for data preparation
and evaluation.