9 papers
STEER: Steerable Dyadic Head Avatars
Kartik Teotia, Helge Rhodin, Hyeongwoo Kim +2
Facial movement and expression are central to face-to-face communication, conveying turn-taking, attention, agreement, and engagement alongside speech. While speech-driven facial a…
Domain Knowledge-Informed Self-Supervised Representations for Workout Form Assessment
Paritosh Parmar, Amol Gharat, Helge Rhodin
Maintaining proper form while exercising is important for preventing injuries and maximizing muscle mass gains. Detecting errors in workout form naturally requires estimating human…
E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation
Mayur Deshmukh, Hiroyasu Akada, Helge Rhodin +2
Event cameras offer multiple advantages in monocular egocentric 3D human pose estimation from head-mounted devices, such as millisecond temporal resolution, high dynamic range, and…
Audio-Driven Universal Gaussian Head Avatars
Kartik Teotia, Helge Rhodin, Mohit Mendiratta +3
We introduce the first method for audio-driven universal photorealistic avatar synthesis, combining a person-agnostic speech model with our novel Universal Head Avatar Prior (UHAP)…
Follow My Hold: Hand-Object Interaction Reconstruction through Geometric Guidance
Ayce Idil Aytekin, Helge Rhodin, Rishabh Dabral +1
We propose a novel diffusion-based framework for reconstructing 3D geometry of hand-held objects from monocular RGB images by leveraging hand-object interaction as geometric guidan…
DreamTexture: Shape from Virtual Texture with Analysis by Augmentation
Ananta R. Bhattarai, Xingzhe He, Alla Sheffer +1
DreamFusion established a new paradigm for unsupervised 3D reconstruction from virtual views by combining advances in generative models and differentiable rendering. However, the u…