5 papers · 1 filter
NGS-Marker: Robust Native Watermarking for 3D Gaussian Splatting
Hao Qin, Yukai Sun, Luyuan Chen +5
With the rapid development and adoption of 3D Gaussian Splatting (3DGS), the need for effective copyright protection has become increasingly critical. Existing watermarking techniq…
SoulX-FlashHead: Oracle-guided Generation of Infinite Real-time Streaming Talking Heads
Tan Yu, Qian Qiao, Le Shen +9
Achieving a balance between high-fidelity visual quality and low-latency streaming remains a formidable challenge in audio-driven portrait generation. Existing large-scale models o…
Distilling Multi-view Diffusion Models into 3D Generators
Hao Qin, Luyuan Chen, Ming Kong +2
We introduce DD3G, a formulation that Distills a multi-view Diffusion model (MV-DM) into a 3D Generator using gaussian splatting. DD3G compresses and integrates extensive visual an…
Enabling Versatile Controls for Video Diffusion Models
Xu Zhang, Hao Zhou, Haoming Qin +5
Despite substantial progress in text-to-video generation, achieving precise and flexible control over fine-grained spatiotemporal attributes remains a significant unresolved challe…
Probablistic Restoration with Adaptive Noise Sampling for 3D Human Pose Estimation
Xianzhou Zeng, Hao Qin, Ming Kong +2
The accuracy and robustness of 3D human pose estimation (HPE) are limited by 2D pose detection errors and 2D to 3D ill-posed challenges, which have drawn great attention to Multi-H…