1 paper
Amir Mann, Gal Michael Harari, Merav Keidar +1
We introduce VideoMDM, a diffusion-based framework that trains 3D human motion priors directly from accurate 2D poses extracted from monocular videos, without any 3D ground truth.…