2 papers
cs.CV2024
Exploring Vision Transformers for 3D Human Motion-Language Models with Motion Patches
Qing Yu, Mikihiro Tanaka, Kent Fujiwara
To build a cross-modal latent space between 3D human motion and language, acquiring large-scale and high-quality human motion data is crucial. However, unlike the abundance of imag…
cs.CV2018
Canonical and Compact Point Cloud Representation for Shape Classification
Kent Fujiwara, Ikuro Sato, Mitsuru Ambai +2
We present a novel compact point cloud representation that is inherently invariant to scale, coordinate change and point permutation. The key idea is to parametrize a distance fiel…