18 papers · 1 filter
Latent Wavelet Diffusion For Ultra-High-Resolution Image Synthesis
Luigi Sigillo, Shengfeng He, Danilo Comminiello
High-resolution image synthesis remains a core challenge in generative modeling, particularly in balancing computational efficiency with the preservation of fine-grained visual det…
MotionAdapter: Video Motion Transfer via Content-Aware Attention Customization
Zhexin Zhang, Yangyang Xu, Yifeng Zhu +4
Recent advances in diffusion-based text-to-video models, particularly those built on the diffusion transformer architecture, have achieved remarkable progress in generating high-qu…
NimbusGS: Unified 3D Scene Reconstruction under Hybrid Weather
Yanying Li, Jinyang Li, Shengfeng He +3
We present NimbusGS, a unified framework for reconstructing high-quality 3D scenes from degraded multi-view inputs captured under diverse and mixed adverse weather conditions. Unli…
OmniVTON++: Training-Free Universal Virtual Try-On with Principal Pose Guidance
Zhaotong Yang, Yong Du, Shengfeng He +5
Image-based Virtual Try-On (VTON) concerns the synthesis of realistic person imagery through garment re-rendering under human pose and body constraints. In practice, however, exist…
Zero-Shot Video Translation via Token Warping
Haiming Zhu, Yangyang Xu, Jun Yu +1
With the revolution of generative AI, video-related tasks have been widely studied. However, current state-of-the-art video models still lag behind image models in visual quality a…
FlashMesh: Faster and Better Autoregressive Mesh Synthesis via Structured Speculation
Tingrui Shen, Yiheng Zhang, Chen Tang +6
Autoregressive models can generate high-quality 3D meshes by sequentially producing vertices and faces, but their token-by-token decoding results in slow inference, limiting practi…