3 papers
cs.CV2024
SpotActor: Training-Free Layout-Controlled Consistent Image Generation
Jiahao Wang, Caixia Yan, Weizhan Zhang +6
Text-to-image diffusion models significantly enhance the efficiency of artistic creation with high-fidelity image generation. However, in typical application scenarios like comic b…
cs.CV2024★ 14 cited
Power-LLaVA: Large Language and Vision Assistant for Power Transmission Line Inspection
Jiahao Wang, Mingxuan Li, Haichen Luo +4
The inspection of power transmission line has achieved notable achievements in the past few years, primarily due to the integration of deep learning technology. However, current in…
cs.CV2022
Towards Real-World Video Deblurring by Exploring Blur Formation Process
Mingdeng Cao, Zhihang Zhong, Yanbo Fan +5
This paper aims at exploring how to synthesize close-to-real blurs that existing video deblurring models trained on them can generalize well to real-world blurry videos. In recent…