5 papers
Context-Aware Autoregressive Diffusion for Gloss-Wise Sign Language Production
JungHoon Sung, Boeun Kim, Chu Xin +4
To generate natural and accurate sentence-level sign language, synthesizing the "gloss", the fundamental semantic unit, is essential. However, most current sign-language production…
Bidirectional Regression for Monocular 6DoF Head Pose Estimation and Reference System Alignment
Sungho Chun, Boeun Kim, Hyung Jin Chang +1
Precise six-degree-of-freedom (6DoF) head pose estimation is crucial for safety-critical applications and human-computer interaction scenarios, yet existing monocular methods still…
High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose Estimation
Runyang Feng, Hyung Jin Chang, Tze Ho Elden Tse +3
Modeling high-resolution spatiotemporal representations, including both global dynamic contexts (e.g., holistic human motion tendencies) and local motion details (e.g., high-freque…
Roll Your Eyes: Gaze Redirection via Explicit 3D Eyeball Rotation
YoungChan Choi, HengFei Wang, YiHua Cheng +4
We propose a novel 3D gaze redirection framework that leverages an explicit 3D eyeball structure. Existing gaze redirection methods are typically based on neural radiance fields, w…
3D Prior is All You Need: Cross-Task Few-shot 2D Gaze Estimation
Yihua Cheng, Hengfei Wang, Zhongqun Zhang +4
3D and 2D gaze estimation share the fundamental objective of capturing eye movements but are traditionally treated as two distinct research domains. In this paper, we introduce a n…