collaborators

7 papers

cs.CV2025

Nested AutoRegressive Models

Hongyu Wu, Xuhui Fan, Zhangkai Wu +1

AutoRegressive (AR) models have demonstrated competitive performance in image generation, achieving results comparable to those of diffusion models. However, their token-by-token i…

cs.CV2025

VALA: Learning Latent Anchors for Training-Free and Temporally Consistent

Zhangkai Wu, Xuhui Fan, Zhongyuan Xie +2

Recent advances in training-free video editing have enabled lightweight and precise cross-frame generation by leveraging pre-trained text-to-image diffusion models. However, existi…

cs.CV2025

FAME: Fairness-aware Attention-modulated Video Editing

Zhangkai Wu, Xuhui Fan, Zhongyuan Xie +3

Training-free video editing (VE) models tend to fall back on gender stereotypes when rendering profession-related prompts. We propose \textbf{FAME} for \textit{Fairness-aware Atten…

cs.CV2025

SCoT: Unifying Consistency Models and Rectified Flows via Straight-Consistent Trajectories

Zhangkai Wu, Xuhui Fan, Hongyu Wu +1

Pre-trained diffusion models are commonly used to generate clean data (e.g., images) from random noises, effectively forming pairs of noises and corresponding clean images. Distill…

cs.LG2025

A Survey on Pre-Trained Diffusion Model Distillations

Xuhui Fan, Zhangkai Wu, Hongyu Wu

Diffusion Models~(DMs) have emerged as the dominant approach in Generative Artificial Intelligence (GenAI), owing to their remarkable performance in tasks such as text-to-image syn…

cs.LG2025

FigBO: A Generalized Acquisition Function Framework with Look-Ahead Capability for Bayesian Optimization

Hui Chen, Xuhui Fan, Zhangkai Wu +1

Bayesian optimization is a powerful technique for optimizing expensive-to-evaluate black-box functions, consisting of two main components: a surrogate model and an acquisition func…