111 citations · 294 across the 8 of their papers we have counts for
1 paper · 1 filter
Fan Bai, Pai Peng, Zhengzhi Tang +8
With the widespread adoption of large multimodal models, efficient inference across text, image, audio, and video modalities has become critical. However, existing multimodal infer…