2 papers
cs.CV2025
StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion
Haoxin Yang, Weihong Chen, Xuemiao Xu +5
Monocular 3D human pose estimation remains a challenging task due to inherent depth ambiguities and occlusions. Compared to traditional methods based on Transformers or Convolution…
cs.CV2025
Action Dubber: Timing Audible Actions via Inflectional Flow
Wenlong Wan, Weiying Zheng, Tianyi Xiang +2
We introduce the task of Audible Action Temporal Localization, which aims to identify the spatio-temporal coordinates of audible movements. Unlike conventional tasks such as action…