activity
20242026
collaborators

8 papers

cs.CV2026

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation

Hongbo Wang, Huaibo Huang, Jie Cao +3

While recent advances in 3D generation have enabled impressive visual synthesis, existing methods often rely on 2D diffusion supervision without explicit mechanisms for geometric c…

cs.CV2026

AnchorSplat: Fast and Structure Consistent Detail Synthesis for Gaussian Splatting

Dexu Zhu, Jiangnan Shao, Xiaofeng Wang +4

3D Gaussian Splatting (3DGS) has emerged as a powerful representation for high-fidelity rendering. However, existing assets often suffer from quality bottlenecks such as missing de…

cs.CV2026

SmartDirector: Keyframe-Conditioned Cinematic Video Generation with Narrative Pacing Control

Zhida Zhang, Jie Ma, Zhan Peng +5

The narrative quality of a video fundamentally determines its perceptual value. Although existing video generation methods can produce visually appealing content, they predominantl…

cs.CV2025

TT-DF: A Large-Scale Diffusion-Based Dataset and Benchmark for Human Body Forgery Detection

Wenkui Yang, Zhida Zhang, Xiaoqiang Zhou +2

The emergence and popularity of facial deepfake methods spur the vigorous development of deepfake datasets and facial forgery detection, which to some extent alleviates the securit…

cs.CV2025

Straighter Flow Matching via a Diffusion-Based Coupling Prior

Siyu Xing, Jie Cao, Huaibo Huang +2

Flow matching as a paradigm of generative model achieves notable success across various domains. However, existing methods use either multi-round training or knowledge within minib…

cs.LG2025

HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling

Hexiong Yang, Mingrui Chen, Huaibo Huang +4

Inspired by the great success of Masked Language Modeling (MLM) in the natural language domain, the paradigm of self-supervised pre-training and fine-tuning has also achieved remar…