5 papers
GSCI: Robust Gaussian Splatting For Snapshot Compressive Imaging via Large Vision Model Priors
Yanming Yang, Chenxi Song, Ping Wang +2
Snapshot Compressive Imaging (SCI) offers an efficient solution for high-speed video acquisition and, under exposure-time camera--scene relative motion, multi-view scene capture by…
Fast3Dcache: Training-free 3D Geometry Synthesis Acceleration
Mengyu Yang, Yanming Yang, Chenyi Xu +5
Diffusion models have achieved impressive generative quality across modalities like 2D images, videos, and 3D shapes, but their inference remains computationally expensive due to t…
Taming Video Models for 3D and 4D Generation via Zero-Shot Camera Control
Chenxi Song, Yanming Yang, Tong Zhao +2
Video diffusion models have rich world priors, but their use in spatial tasks is limited by poor control, spatial-temporal inconsistent results, and entangled scene-camera dynamics…
FlowDirector: Training-Free Flow Steering for Precise Text-to-Video Editing
Guangzhao Li, Yanming Yang, Chenxi Song +1
Text-driven video editing aims to modify video content based on natural language instructions. While recent training-free methods have leveraged pretrained diffusion models, they o…
Distill Any Depth: Distillation Creates a Stronger Monocular Depth Estimator
Xiankang He, Dongyan Guo, Hongji Li +3
Recent advances in zero-shot monocular depth estimation(MDE) have significantly improved generalization by unifying depth distributions through normalized depth representations and…