2 papers
cs.HC2025
SimVecVis: A Dataset for Enhancing MLLMs in Visualization Understanding
Can Liu, Chunlin Da, Xiaoxiao Long +3
Current multimodal large language models (MLLMs), while effective in natural image understanding, struggle with visualization understanding due to their inability to decode the dat…
cs.CV2025
DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation
Junhao Chen, Mingjin Chen, Jianjin Xu +9
Controllable video generation (CVG) has advanced rapidly, yet current systems falter when more than one actor must move, interact, and exchange positions under noisy control signal…