Story
arxiv_cs_ai ยท Aug 27, 2026 ยท paper
arxiv.orgAug 27, 2026
original source linked
In brief
Vision-Language Models (VLMs) demonstrate exceptional visual reasoning capabilities, yet their inference costs escalate rapidly with the proliferation of visual tokens. Existing visual token pruning methods exhibit tw...
Feed lens
eval