9 citations · 17 across the 3 of their papers we have counts for
5 papers · 1 filter
Reconstructing Action-Conditioned Human-Object Interactions Using Commonsense Knowledge Priors
Xi Wang, Gen Li, Yen-Ling Kuo +3
We present a method for inferring diverse 3D models of human-object interactions from images. Reasoning about how humans interact with objects in complex scenes from a single 2D im…
LiP-Flow: Learning Inference-time Priors for Codec Avatars via Normalizing Flows in Latent Space
Emre Aksan, Shugao Ma, Akin Caliskan +5
Neural face avatars that are trained from multi-view data captured in camera domes can produce photo-realistic 3D reconstructions. However, at inference time, they must be driven b…
Towards End-to-end Video-based Eye-Tracking
Seonwook Park, Emre Aksan, Xucong Zhang +1
Estimating eye-gaze from images alone is a challenging task, in large parts due to un-observable person-specific factors. Achieving high accuracy typically requires labeled data fr…
Structured Prediction Helps 3D Human Motion Modelling
Emre Aksan, Manuel Kaufmann, Otmar Hilliges
Human motion prediction is a challenging and important task in many computer vision application domains. Existing work only implicitly models the spatial structure of the human ske…
Guiding InfoGAN with Semi-Supervision
Adrian Spurr, Emre Aksan, Otmar Hilliges
In this paper we propose a new semi-supervised GAN architecture (ss-InfoGAN) for image synthesis that leverages information from few labels (as little as 0.22%, max. 10% of the dat…