Preprint
Aug 2026
PACE: A Unified Condense-and-Extract Paradigm for Fast VLM Inference
PACE (Pixel-Adaptive Condense and Extract), a training-free inference framework that accelerates both the vision encoder and the Large Language Model (LLM) via a unified Condense-and-Extract paradigm, is proposed.
Jun-Jie Liu, Shengyuan Ye, Xu Chen
· 0 citations