6 papers
KV-Control: Parameter-Efficient K/V Injection for Trajectory-Controlled Text-to-Motion
Tengjiao Sun, Pengcheng Fang, Xiaoyu Zhan +4
Text-conditioned 3D human motion models now synthesize plausible motions from prompts, but practical animation and embodied-agent workflows rarely stop at text: a character may nee…
MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning
Yi-Yang Zhang, Tengjiao Sun, Pengcheng Fang +4
3D Human motion generation is pivotal across film, animation, gaming, and embodied intelligence. Traditional 3D motion synthesis relies on costly motion capture, while recent work…
AnchorRoute: Human Motion Synthesis with Interval-Routed Sparse Contro
Pengcheng Fang, Tengjiao Sun, Dongjie Fu +4
Sparse anchors provide a compact interface for human motion authoring: users specify a few root positions, planar trajectory samples, or body-point targets, while the system synthe…
UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars
Xiaoyu Zhan, Xinyu Fu, Chenghao Yang +9
Speech-driven gestures and facial animations are fundamental to expressive digital avatars in games, virtual production, and interactive media. However, existing methods are either…
SHIELD: Scalable Optimal Control with Certification using Duality and Convexity
Hansung Kim, Siddharth H. Nair, Francesco Borrelli
We present SHIELD, a hierarchical algorithm that reduces both the decision-variable dimension and the constraint set in -regularized convex programs. From strong convexity…
MOGO: Residual Quantized Hierarchical Causal Transformer for High-Quality and Real-Time 3D Human Motion Generation
Dongjie Fu, Tengjiao Sun, Pengcheng Fang +2
Recent advances in transformer-based text-to-motion generation have led to impressive progress in synthesizing high-quality human motion. Nevertheless, jointly achieving high fidel…