16 citations · 16 across the 2 of their papers we have counts for
1 paper · 1 filter
Morunliu Yang, Ruotao Xu, Le Li +6
Omnimodal large language models (OmniLLMs) have recently gained increasing attention for unified audio-video understanding. However, processing long multimodal token sequences intr…