3 papers
cs.CV2026
Where a New Concept Must Enter: Entry Point Gates Cross-Task Usability in Unified Multimodal Models
Zongyang Qiu, Yihan Wu, Kaixuan Fan +2
Unified multimodal models (UMMs) are motivated by the hope that understanding and generation reinforce each other but controlled ablations repeatedly find that adding a generation…
cs.CV2026
EmoSpace: Immersive Affective Image Generation Guided by Fine-Grained Emotion Prototypes
Bingyuan Wang, Xingbei Chen, Zongyang Qiu +2
Immersive affective content generation aims to create visually compelling VR imagery with controllable emotional nuance, yet existing methods typically rely on coarse labels or pro…
cs.CV2025
EmoVid: A Multimodal Emotion Video Dataset for Emotion-Centric Video Understanding and Generation
Zongyang Qiu, Bingyuan Wang, Xingbei Chen +2
Emotion plays a pivotal role in video-based expression, but existing video generation systems predominantly focus on low-level visual metrics while neglecting affective dimensions.…