2 papers
cs.AI2026
AffectOmni: RL-Verifiable People-Centric Grounded Affective Reasoning for Social and Art-Related Scenes
Yibo Wang, Rui Yang, Jisheng Dang +7
Multimodal large language models (MLLMs) achieve strong performance on VQA and scene understanding, yet affective reasoning remains vulnerable to shortcut behavior. Models may pred…
cs.CV2024
AuthFormer: Adaptive Multimodal biometric authentication transformer for middle-aged and elderly people
Yang rui, Meng ling-tao, Zhang qiu-yu
Multimodal biometric authentication methods address the limitations of unimodal biometric technologies in security, robustness, and user adaptability. However, most existing method…