3 papers
cs.AI2026
UniMo: Unified Motion Generation and Understanding with Chain of Thought
Guocun Wang, Kenkun Liu, Jing Lin +3
Existing 3D human motion generation and understanding methods often exhibit limited interpretability, restricting effective mutual enhancement between these inherently related task…
cs.CV2025
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
Yuxiao Yang, Hualian Sheng, Sijia Cai +6
Video generation models have advanced significantly, yet they still struggle to synthesize complex human movements due to the high degrees of freedom in human articulation. This li…
cs.CV2025
Towards Fine-Grained Human Motion Video Captioning
Guorui Song, Guocun Wang, Zhe Huang +4
Generating accurate descriptions of human actions in videos remains a challenging task for video captioning models. Existing approaches often struggle to capture fine-grained motio…