4 papers
Fine-grained Human Motion Understanding with Language Models
Thomas Markhorst, Zhi-Yi Lin, Jouh Yeong Chew +2
In this work, we propose \methodname, an LLM-based model for fine-grained human motion understanding that represents motion as a sequence of skeletal poses with explicit timestamps…
BridgeDiff: Bridging Human Observations and Flat-Garment Synthesis for Virtual Try-Off
Shuang Liu, Ao Yu, Linkang Cheng +5
Virtual try-off (VTOFF) aims to recover canonical flat-garment representations from images of dressed persons for standardized display and downstream virtual try-on. Prior methods…
GazeHTA: End-to-end Gaze Target Detection with Head-Target Association
Zhi-Yi Lin, Jouh Yeong Chew, Jan van Gemert +1
Precisely detecting which object a person is paying attention to is critical for human-robot interaction since it provides important cues for the next action from the human user. W…
3D Kinematics Estimation from Video with a Biomechanical Model and Synthetic Training Data
Zhi-Yi Lin, Bofan Lyu, Judith Cueto Fernandez +3
Accurate 3D kinematics estimation of human body is crucial in various applications for human health and mobility, such as rehabilitation, injury prevention, and diagnosis, as it he…