← Back to Model Beat
Research·Jul 7·all news from July 7, 2026

LensVLM: Selective Context Expansion for Compressed Visual Representation of Text

Apple researchers have introduced LensVLM, a method that allows vision language models to process text as images rather than traditional token sequences. By utilizing selective context expansion, this approach maintains efficiency by adjusting rendering resolution, which potentially improves how models handle long-form text while bypassing standard tokenization limitations.

Covered by 1 source

Related stories

ResearchIncentivizing Temporal-Awareness in Egocentric Video Understanding ModelsJul 7 · 6 sourcesResearchOpenAI may have made a fatal misstep in copyright fight with news orgsJul 9 · 6 sourcesResearchOpenAI finds roughly 30 percent of popular AI coding test is brokenJul 9ResearchRaytheon, Rheinmetall Anchor AI Training Effort for UK ArmyJul 10