1 paper
Zheyu Huang, Zijing Shi, Haozhe Luo +4
Recent advances in Large Multimodal Models (LMMs) have greatly improved video understanding, yet their ability to reason about human-centered social situations remains limited. Exi…