8 papers
Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation
Hongbo Wang, Huaibo Huang, Jie Cao +3
While recent advances in 3D generation have enabled impressive visual synthesis, existing methods often rely on 2D diffusion supervision without explicit mechanisms for geometric c…
AnchorSplat: Fast and Structure Consistent Detail Synthesis for Gaussian Splatting
Dexu Zhu, Jiangnan Shao, Xiaofeng Wang +4
3D Gaussian Splatting (3DGS) has emerged as a powerful representation for high-fidelity rendering. However, existing assets often suffer from quality bottlenecks such as missing de…
SmartDirector: Keyframe-Conditioned Cinematic Video Generation with Narrative Pacing Control
Zhida Zhang, Jie Ma, Zhan Peng +5
The narrative quality of a video fundamentally determines its perceptual value. Although existing video generation methods can produce visually appealing content, they predominantl…
TT-DF: A Large-Scale Diffusion-Based Dataset and Benchmark for Human Body Forgery Detection
Wenkui Yang, Zhida Zhang, Xiaoqiang Zhou +2
The emergence and popularity of facial deepfake methods spur the vigorous development of deepfake datasets and facial forgery detection, which to some extent alleviates the securit…
Straighter Flow Matching via a Diffusion-Based Coupling Prior
Siyu Xing, Jie Cao, Huaibo Huang +2
Flow matching as a paradigm of generative model achieves notable success across various domains. However, existing methods use either multi-round training or knowledge within minib…
HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling
Hexiong Yang, Mingrui Chen, Huaibo Huang +4
Inspired by the great success of Masked Language Modeling (MLM) in the natural language domain, the paradigm of self-supervised pre-training and fine-tuning has also achieved remar…