11 papers
SHIFT: Motion Alignment in Video Diffusion Models with Adversarial Hybrid Fine-Tuning
Xi Ye, Wenjia Yang, Yangyang Xu +4
Image-conditioned video diffusion models achieve impressive visual realism but often suffer from weakened motion fidelity, e.g., reduced motion dynamics or degraded long-term tempo…
MotionAdapter: Video Motion Transfer via Content-Aware Attention Customization
Zhexin Zhang, Yangyang Xu, Yifeng Zhu +4
Recent advances in diffusion-based text-to-video models, particularly those built on the diffusion transformer architecture, have achieved remarkable progress in generating high-qu…
NimbusGS: Unified 3D Scene Reconstruction under Hybrid Weather
Yanying Li, Jinyang Li, Shengfeng He +3
We present NimbusGS, a unified framework for reconstructing high-quality 3D scenes from degraded multi-view inputs captured under diverse and mixed adverse weather conditions. Unli…
OmniVTON++: Training-Free Universal Virtual Try-On with Principal Pose Guidance
Zhaotong Yang, Yong Du, Shengfeng He +5
Image-based Virtual Try-On (VTON) concerns the synthesis of realistic person imagery through garment re-rendering under human pose and body constraints. In practice, however, exist…
Zero-Shot Video Translation via Token Warping
Haiming Zhu, Yangyang Xu, Jun Yu +1
With the revolution of generative AI, video-related tasks have been widely studied. However, current state-of-the-art video models still lag behind image models in visual quality a…
DeshadowMamba: Deshadowing as 1D Sequential Similarity
Zhaotong Yang, Yi Chen, Yanying Li +5
Recent deep models for image shadow removal often rely on attention-based architectures to capture long-range dependencies. However, their fixed attention patterns tend to mix illu…