3 papers
cs.AI2026
Debiased Multimodal Personality Understanding through Dual Causal Intervention
Yangfu Zhu, Zitong Han, Nianwen Ning +4
Multimodalpersonalityunderstandingplaysacriticalroleinhuman centered artificial intelligence. Previous work mainly focus on learn-ing rich multimodal representations for video pers…
cs.CV2025
MathSight: A Benchmark Exploring Have Vision-Language Models Really Seen in University-Level Mathematical Reasoning?
Yuandong Wang, Yao Cui, Yuxin Zhao +3
Recent advances in Vision-Language Models (VLMs) have achieved impressive progress in multimodal mathematical reasoning. Yet, how much visual information truly contributes to reaso…
cs.RO2025
Eq.Bot: Enhance Robotic Manipulation Learning via Group Equivariant Canonicalization
Jian Deng, Yuandong Wang, Yangfu Zhu +3
Robotic manipulation systems are increasingly deployed across diverse domains. Yet existing multi-modal learning frameworks lack inherent guarantees of geometric consistency, strug…