3 papers
cs.CV2026
PRISM: Streaming Human Motion Generation with Per-Joint Latent Decomposition
Zeyu Ling, Qing Shuai, Teng Zhang +3
Text-to-motion generation has advanced with larger corpora and stronger generators, yet many models still rely on holistic frame- or clip-level latents that entangle trajectory, or…
cs.GR2025
Diff-3DCap: Shape Captioning with Diffusion Models
Zhenyu Shu, Jiawei Wen, Shiyang Li +2
The task of 3D shape captioning occupies a significant place within the domain of computer graphics and has garnered considerable interest in recent years. Traditional approaches t…
cs.GR2025
Sketch3DVE: Sketch-based 3D-Aware Scene Video Editing
Feng-Lin Liu, Shi-Yang Li, Yan-Pei Cao +2
Recent video editing methods achieve attractive results in style transfer or appearance modification. However, editing the structural content of 3D scenes in videos remains challen…