3 papers
cs.CV2025
JCo-MVTON: Jointly Controllable Multi-Modal Diffusion Transformer for Mask-Free Virtual Try-on
Aowen Wang, Wei Li, Hao Luo +4
Virtual try-on systems have long been hindered by heavy reliance on human body masks, limited fine-grained control over garment attributes, and poor generalization to real-world, i…
cs.CV2024
Coherent Video Inpainting Using Optical Flow-Guided Efficient Diffusion
Bohai Gu, Hao Luo, Song Guo +2
The text-guided video inpainting technique has significantly improved the performance of content generation applications. A recent family for these improvements uses diffusion mode…
cs.CV2024
Fisheye-GS: Lightweight and Extensible Gaussian Splatting Module for Fisheye Cameras
Zimu Liao, Siyan Chen, Rong Fu +9
Recently, 3D Gaussian Splatting (3DGS) has garnered attention for its high fidelity and real-time rendering. However, adapting 3DGS to different camera models, particularly fisheye…