Tech News
← Home  ·  All topics

Long Context Compression

1 GoKawiil brief on this topic

Apple releases LensVLM-9B, a vision-language model that compresses text as images

Apple's AI research team has published LensVLM-9B, a 9-billion-parameter vision language model that processes long documents by first compressing text into image form and then selectively decompressing only the pages it judges relevant, using learned tools. The model, code and demo scripts have been released on GitHub and Hugging Face under Apple's Machine Learning Research Model License and Sample Code License, with configurable compression ratios of 5x, 10x, or 15x.