1 paper · 1 filter
Zhenghui Guo, Yilin Yang, Yuanbin Man +5
Token compression in OmniLLMs is typically posed as a single saliency-ranking problem: score each multimodal token, keep the top-K. We argue this abstraction is mis-specified. The…