2 papers
cs.CV2025
LUMA: Low-Dimension Unified Motion Alignment with Dual-Path Anchoring for Text-to-Motion Diffusion Model
Haozhe Jia, Wenshuo Chen, Yuqi Lin +8
While current diffusion-based models, typically built on U-Net architectures, have shown promising results on the text-to-motion generation task, they still suffer from semantic mi…
cs.CV2024
When Spatial meets Temporal in Action Recognition
Huilin Chen, Lei Wang, Yifan Chen +2
Video action recognition has made significant strides, but challenges remain in effectively using both spatial and temporal information. While existing methods often focus on eithe…