arXiv Artificial Intelligence

VPRune: Efficient Training-free Pre-LLM Visual Token Pruning

VPRune: Efficient Training-free Pre-LLM Visual Token Pruning

Quick summary

arXiv:2609.24485v2 Announce Type: cross Abstract: Visual token pruning is a promising approach to reducing the inference cost of large vision-language models (LVLMs), yet aggressive token reduction often causes substantial performance degradation. We identify three key factors behind this degradation: text-guided selection bias, information loss from discarded tokens, and positional distortion caused by sequence compaction. Based on these observations, we propose \textbf{VPRune}, a training-free pre-LLM pruning framework consisting of visual-only diversity selection, similarity-guided token re

Key takeaways

  • arXiv:2609.24485v2 Announce Type: cross Abstract: Visual token pruning is a promising approach to reducing the inference cost of large vision-language models (LVLMs), yet aggressive token reduction often causes substantial performance degradation.
  • We identify three key factors behind this degradation: text-guided selection bias, information loss from discarded tokens, and positional distortion caused by sequence compaction.
  • Based on these observations, we propose \textbf{VPRune}, a training-free pre-LLM pruning framework consisting of visual-only diversity selection, similarity-guided token re

Why it matters

The importance of “VPRune: Efficient Training-free Pre-LLM Visual Token Pruning” will be measured by what changes in practice. User behavior, access conditions, verifiable performance and responsible-use outcomes are the signals worth following.

Kaynak sitede devamını oku: arXiv Artificial Intelligence ↗