3 papers
cs.CV2025
HAT: Hybrid Attention Transformer for Image Restoration
Xiangyu Chen, Xintao Wang, Wenlong Zhang +4
Transformer-based methods have shown impressive performance in image restoration tasks, such as image super-resolution and denoising. However, we find that these networks can only…
cs.CV2025
Exploring Scalable Unified Modeling for General Low-Level Vision
Xiangyu Chen, Kaiwen Zhu, Yuandong Pu +7
Low-level vision involves a wide spectrum of tasks, including image restoration, enhancement, stylization, and feature extraction, which differ significantly in both task formulati…
cs.CV2025
Lumina-OmniLV: A Unified Multimodal Framework for General Low-Level Vision
Yuandong Pu, Le Zhuo, Kaiwen Zhu +7
We present Lunima-OmniLV (abbreviated as OmniLV), a universal multimodal multi-task framework for low-level vision that addresses over 100 sub-tasks across four major categories: i…