1 paper
Qing Yu, Mikihiro Tanaka, Kent Fujiwara
To build a cross-modal latent space between 3D human motion and language, acquiring large-scale and high-quality human motion data is crucial. However, unlike the abundance of imag…