cs.AI 2605.05668

Large Vision-Language Models Get Lost in Attention

Using an information-theoretic and geometric framework, the paper reveals that attention mainly reconfigures representations while FFN drives semantic innovation, exposing redundancy in LVLMs.

Gongli Xi, Ye Tian, Mengyu Yang et al.

2026-05-07 49