Index-Preserving Lightweight Token Pruning for Efficient Document Understanding in Vision-Language Models (ICLR 2026 Workshop on MM Intelligence, Poster).
vlmmultimodaldocument-understandinginference-accelerationtoken-pruningvision-language-modelsefficient-aiiclr-2026patch-pruning
-
Updated
Jun 11, 2026 - Python