1 paper
Fuyu Xing, Zimu Wang, Wei Wang +1
The proliferation of multimedia content necessitates the development of effective Multimedia Event Extraction (M2E2) systems. Though Large Vision-Language Models (LVLMs) have shown…