4 papers · 1 filter
UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception
Qin Guo, Hao Luo, Dongxu Yue +4
Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing approaches typically treat diff…
JCo-MVTON: Jointly Controllable Multi-Modal Diffusion Transformer for Mask-Free Virtual Try-on
Aowen Wang, Wei Li, Hao Luo +4
Virtual try-on systems have long been hindered by heavy reliance on human body masks, limited fine-grained control over garment attributes, and poor generalization to real-world, i…
Coherent Video Inpainting Using Optical Flow-Guided Efficient Diffusion
Bohai Gu, Hao Luo, Song Guo +2
The text-guided video inpainting technique has significantly improved the performance of content generation applications. A recent family for these improvements uses diffusion mode…
Fisheye-GS: Lightweight and Extensible Gaussian Splatting Module for Fisheye Cameras
Zimu Liao, Siyan Chen, Rong Fu +9
Recently, 3D Gaussian Splatting (3DGS) has garnered attention for its high fidelity and real-time rendering. However, adapting 3DGS to different camera models, particularly fisheye…