multimodal
fact
bullish
PACE framework enables Qwen2.5-VL-7B to retain 93.8% of original performance while using only 10% of visual tokens, achieving 3.1x speedup in time to first token
By integrating PACE into Qwen2.5-VL-7B, the model retains 93.8% of its original performance while utilizing only 10% of the visual tokens, yielding a 3.1x speedup in time to first token (TTFT).
Computer Vision30 Aug 2026