5 papers
Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning
Yu Gu, Zijun Yu, Vahid Partovi Nia +1
Chain-of-thought (CoT) reasoning with self-consistency improves performance by aggregating multiple sampled reasoning paths. In this setting, correctness is no longer tied to a sin…
AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation
Yuchao Gu, Guian Fang, Yuxin Jiang +4
Few-step video generation has been significantly advanced by consistency distillation. However, the performance of consistency-distilled models often degrades as more sampling step…
EditTransfer++: Toward Faithful and Efficient Visual-Prompt-Guided Image Editing
Lan Chen, Qi Mao, Yiren Song +2
Visual-prompt-guided edit transfer aims to learn image transformations directly from example pairs, offering more precise and controllable editing than purely text-driven approache…
DC-Gen: Post-Training Diffusion Acceleration with Deeply Compressed Latent Space
Wenkun He, Yuchao Gu, Junyu Chen +11
Existing text-to-image diffusion models excel at generating high-quality images, but face significant efficiency challenges when scaled to high resolutions, like 4K image generatio…
DC-VideoGen: Efficient Video Generation with Deep Compression Video Autoencoder
Junyu Chen, Wenkun He, Yuchao Gu +12
We introduce DC-VideoGen, a post-training acceleration framework for efficient video generation. DC-VideoGen can be applied to any pre-trained video diffusion model, improving effi…