6 papers · 1 filter
Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars
Yicheng Gong, Jiawei Zhang, Liqiang Liu +6
We present a framework for explicit emotion control in feed-forward, single-image 3D head avatar reconstruction. Unlike existing pipelines where emotion is implicitly entangled wit…
Mango-GS: Enhancing Spatio-Temporal Consistency in Dynamic Scenes Reconstruction using Multi-Frame Node-Guided 4D Gaussian Splatting
Tingxuan Huang, Haowei Zhu, Jun-hai Yong +2
Reconstructing dynamic 3D scenes with photorealistic detail and strong temporal coherence remains a significant challenge. Existing Gaussian splatting approaches for dynamic scene…
Dynamic Gaussian Scene Reconstruction from Unsynchronized Videos
Zhixin Xu, Hengyu Zhou, Yuan Liu +4
Multi-view video reconstruction plays a vital role in computer vision, enabling applications in film production, virtual reality, and motion analysis. While recent advances such as…
StructRe: Rewriting for Structured Shape Modeling
Jiepeng Wang, Hao Pan, Yang Liu +3
Man-made 3D shapes are naturally organized in parts and hierarchies; such structures provide important constraints for shape reconstruction and generation. Modeling shape structure…
MMGen: Unified Multi-modal Image Generation and Understanding in One Go
Jiepeng Wang, Zhaoqing Wang, Hao Pan +4
A unified diffusion framework for multi-modal generation and understanding has the transformative potential to achieve seamless and controllable image diffusion and other cross-mod…
Generative Hierarchical Temporal Transformer for Hand Pose and Action Modeling
Yilin Wen, Hao Pan, Takehiko Ohkawa +5
We present a novel unified framework that concurrently tackles recognition and future prediction for human hand pose and action modeling. Previous works generally provide isolated…