HF Daily Papers
· Papers
OmniScope: Modality-Decoupled Token Compression for Omnimodal Large Language Models
Existing token compression methods for omnimodal large language models typically rely on one modality to determine what to retain in the other. We show that this assumption often breaks down: for the same query, audio and video relevance often peaks at different moments. This cross-modal salience mi