6 papers
Bilingual Text-to-Motion Generation: A New Benchmark and Baselines
Wanjiang Weng, Xiaofeng Tan, Xiangbo Shu +3
Text-to-motion generation holds significant potential for cross-linguistic applications, yet it is hindered by the lack of bilingual datasets and the poor cross-lingual semantic un…
ReAlign: Text-to-Motion Generation via Step-Aware Reward-Guided Alignment
Wanjiang Weng, Xiaofeng Tan, Junbo Wang +3
Text-to-motion generation, which synthesizes 3D human motions from text inputs, holds immense potential for applications in gaming, film, and robotics. Recently, diffusion-based me…
Foundation Model for Skeleton-Based Human Action Understanding
Hongsong Wang, Wanjiang Weng, Junbo Wang +4
Human action understanding serves as a foundational pillar in the field of intelligent motion perception. Skeletons serve as a modality- and device-agnostic representation for huma…
PAMD: Plausibility-Aware Motion Diffusion Model for Long Dance Generation
Hongsong Wang, Yin Zhu, Qiuxia Lai +3
Computational dance generation is crucial in many areas, such as art, human-computer interaction, virtual reality, and digital entertainment, particularly for generating coherent a…
SAM-Aware Graph Prompt Reasoning Network for Cross-Domain Few-Shot Segmentation
Shi-Feng Peng, Guolei Sun, Yong Li +2
The primary challenge of cross-domain few-shot segmentation (CD-FSS) is the domain disparity between the training and inference phases, which can exist in either the input data or…
USDRL: Unified Skeleton-Based Dense Representation Learning with Multi-Grained Feature Decorrelation
Wanjiang Weng, Hongsong Wang, Junbo Wang +2
Contrastive learning has achieved great success in skeleton-based representation learning recently. However, the prevailing methods are predominantly negative-based, necessitating…