1 paper
Tianchen Guo, Chen Liu, Ling Chen +1
Multimodal Large Language Models (MLLMs) have shown remarkable progress in single-image perception, yet their ability to reason about complex cross-view human-centric scenes remain…