most citedTaming Video Models for 3D and 4D Generation via Zero-Shot Camera Control

1 citations · 1 across the 6 of their papers we have counts for

collaborators

8 papers

cs.CV2026

BAG: Budget-Aware Gating for Diffusion Caching

Tong Zhao, Mingkun Lei, Yucheng Han +1

Diffusion caching is a lightweight strategy that accelerates Diffusion Transformers (DiTs) by reusing intermediate features across denoising steps, but existing paradigms face a fu…

cs.CV2026

Budget-Constrained Step-Level Diffusion Caching

Mingkun Lei, Tong Zhao, Liangyu Yuan +1

Step-level caching accelerates diffusion models by exploiting temporal redundancy across denoising steps. Existing methods make per-step cache decisions using threshold-based heuri…

cs.CV2026

Fast3Dcache: Training-free 3D Geometry Synthesis Acceleration

Mengyu Yang, Yanming Yang, Chenyi Xu +5

Diffusion models have achieved impressive generative quality across modalities like 2D images, videos, and 3D shapes, but their inference remains computationally expensive due to t…

cs.CV2026

Improving Diffusion Generalization with Weak-to-Strong Segmented Guidance

Liangyu Yuan, Yufei Huang, Mingkun Lei +5

Diffusion models generate synthetic images through an iterative refinement process. However, the misalignment between the simulation-free objective and the iterative process often…

cs.GR20261 cited

Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control

Chenxi Song, Yanming Yang, Tong Zhao +2

Video diffusion models have rich world priors, but their use in spatial tasks is limited by poor control, spatial-temporal inconsistent results, and entangled scene-camera dynamics…

cs.CV2026

Few-Step Diffusion Sampling Through Instance-Aware Discretizations

Liangyu Yuan, Ruoyu Wang, Tong Zhao +4

Diffusion and flow matching models generate high-fidelity data by simulating paths defined by Ordinary or Stochastic Differential Equations (ODEs/SDEs), starting from a tractable p…