Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
AffectOmni: RL-Verifiable People-Centric Grounded Affective Reasoning for Social and Art-Related Scenes
Yibo Wang, Rui Yang, Jisheng Dang +7
Multimodal large language models (MLLMs) achieve strong performance on VQA and scene understanding, yet affective reasoning remains vulnerable to shortcut behavior. Models may pred…
cs.AI2024
A Picture Is Worth a Graph: A Blueprint Debate Paradigm for Multimodal Reasoning
Changmeng Zheng, Dayong Liang, Wengyu Zhang +3
This paper presents a pilot study aimed at introducing multi-agent debate into multimodal reasoning. The study addresses two key challenges: the trivialization of opinions resultin…