4 papers
UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation
Ziyi Wang, Xinshun Wang, Shuang Chen +2
We present UniMotion, to our knowledge the first unified framework for simultaneous understanding and generation of human motion, natural language, and RGB images within a single a…
Crafting Your Evolving Dreams: Concept-Incremental Versatile Customization
Jiahua Dong, Wenqi Liang, Hongliu Li +7
Custom diffusion models (CDMs) have garnered significant interest owing to their remarkable capacity for generating personalized concepts. However, the majority of CDMs unrealistic…
Never-Ending Behavior-Cloning Agent for Robotic Manipulation
Wenqi Liang, Gan Sun, Yao He +3
Relying on multi-modal observations, embodied robots (e.g., humanoid robots) could perform multiple robotic manipulation tasks in unstructured real-world environments. However, mos…
Domain Consistency Representation Learning for Lifelong Person Re-Identification
Shiben Liu, Huijie Fan, Qiang Wang +3
Lifelong person re-identification (LReID) exhibits a contradictory relationship between intra-domain discrimination and inter-domain gaps when learning from continuous data. Intra-…