4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2024
Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation
Peng Jin, Hao Li, Zesen Cheng +6
Text-to-motion generation requires not only grounding local actions in language but also seamlessly blending these individual actions to synthesize diverse and realistic global mot…
cs.CV2022★ 4 cited
Locality Guidance for Improving Vision Transformers on Tiny Datasets
Kehan Li, Runyi Yu, Zhennan Wang +3
While the Vision Transformer (VT) architecture is becoming trendy in computer vision, pure VT models perform poorly on tiny datasets. To address this issue, this paper proposes the…