activity
20242026
collaborators

9 papers

cs.CV2026

STEER: Steerable Dyadic Head Avatars

Kartik Teotia, Helge Rhodin, Hyeongwoo Kim +2

Facial movement and expression are central to face-to-face communication, conveying turn-taking, attention, agreement, and engagement alongside speech. While speech-driven facial a…

cs.CV2026

Domain Knowledge-Informed Self-Supervised Representations for Workout Form Assessment

Paritosh Parmar, Amol Gharat, Helge Rhodin

Maintaining proper form while exercising is important for preventing injuries and maximizing muscle mass gains. Detecting errors in workout form naturally requires estimating human…

cs.CV2026

E-3DPSM: A State Machine for Event-Based Egocentric 3D Human Pose Estimation

Mayur Deshmukh, Hiroyasu Akada, Helge Rhodin +2

Event cameras offer multiple advantages in monocular egocentric 3D human pose estimation from head-mounted devices, such as millisecond temporal resolution, high dynamic range, and…

cs.CV2025

Audio-Driven Universal Gaussian Head Avatars

Kartik Teotia, Helge Rhodin, Mohit Mendiratta +3

We introduce the first method for audio-driven universal photorealistic avatar synthesis, combining a person-agnostic speech model with our novel Universal Head Avatar Prior (UHAP)…

cs.CV2025

Follow My Hold: Hand-Object Interaction Reconstruction through Geometric Guidance

Ayce Idil Aytekin, Helge Rhodin, Rishabh Dabral +1

We propose a novel diffusion-based framework for reconstructing 3D geometry of hand-held objects from monocular RGB images by leveraging hand-object interaction as geometric guidan…

cs.CV2025

DreamTexture: Shape from Virtual Texture with Analysis by Augmentation

Ananta R. Bhattarai, Xingzhe He, Alla Sheffer +1

DreamFusion established a new paradigm for unsupervised 3D reconstruction from virtual views by combining advances in generative models and differentiable rendering. However, the u…