13 citations · 16 across the 6 of their papers we have counts for
6 papers
Lite2Relight: 3D-aware Single Image Portrait Relighting
Pramod Rao, Gereon Fox, Abhimitra Meka +8
Achieving photorealistic 3D view synthesis and relighting of human portraits is pivotal for advancing AR/VR applications. Existing methodologies in portrait relighting demonstrate…
Live2Diff: Live Stream Translation via Uni-directional Attention in Video Diffusion Models
Zhening Xing, Gereon Fox, Yanhong Zeng +4
Large Language Models have shown remarkable efficacy in generating streaming data such as text and audio, thanks to their temporally uni-directional attention mechanism, which mode…
AvatarStudio: Text-driven Editing of 3D Dynamic Human Head Avatars
Mohit Mendiratta, Xingang Pan, Mohamed Elgharib +6
Capturing and editing full head performances enables the creation of virtual characters with various applications such as extended reality and media production. The past few years…
GVP: Generative Volumetric Primitives
Mallikarjun B R, Xingang Pan, Mohamed Elgharib +1
Advances in 3D-aware generative models have pushed the boundary of image synthesis with explicit camera control. To achieve high-resolution image synthesis, several attempts have b…
HQ3DAvatar: High Quality Controllable 3D Head Avatar
Kartik Teotia, Mallikarjun B R, Xingang Pan +4
Multi-view volumetric rendering techniques have recently shown great potential in modeling and synthesizing high-quality head avatars. A common approach to capture full head dynami…
Retrieval in Long Surveillance Videos using User Described Motion and Object Attributes
Greg Castanon, Mohamed Elgharib, Venkatesh Saligrama +1
We present a content-based retrieval method for long surveillance videos both for wide-area (Airborne) as well as near-field imagery (CCTV). Our goal is to retrieve video segments,…