Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Hierarchical Motion Captioning Utilizing External Text Data Source
Clayton Leite, Yu Xiao
This paper introduces a novel approach to enhance existing motion captioning methods, which directly map representations of movement to high-level descriptive captions (e.g., ``a p…
cs.LG2024
Transformer-Based Approaches for Sensor-Based Human Activity Recognition: Opportunities and Challenges
Clayton Souza Leite, Henry Mauranen, Aziza Zhanabatyrova +1
Transformers have excelled in natural language processing and computer vision, paving their way to sensor-based Human Activity Recognition (HAR). Previous studies show that transfo…
cs.LG2024
Enhancing Motion Variation in Text-to-Motion Models via Pose and Video Conditioned Editing
Clayton Leite, Yu Xiao
Text-to-motion models that generate sequences of human poses from textual descriptions are garnering significant attention. However, due to data scarcity, the range of motions thes…