3 citations · 3 across the 5 of their papers we have counts for
1 paper · 1 filter
Fan Bai, Pai Peng, Zhengzhi Tang +8
With the widespread adoption of large multimodal models, efficient inference across text, image, audio, and video modalities has become critical. However, existing multimodal infer…