3 papers
cs.CV2026
Seg-Agent: Test-Time Multimodal Reasoning for Training-Free Language-Guided Segmentation
Chao Hao, Jun Xu, Ji Du +6
Language-guided segmentation transcends the scope limitations of traditional semantic segmentation, enabling models to segment arbitrary target regions based on natural language in…
cs.CV2024
Dilated Strip Attention Network for Image Restoration
Fangwei Hao, Jiesheng Wu, Ji Du +2
Image restoration is a long-standing task that seeks to recover the latent sharp image from its deteriorated counterpart. Due to the robust capacity of self-attention to capture lo…
eess.IV2024
Large coordinate kernel attention network for lightweight image super-resolution
Fangwei Hao, Jiesheng Wu, Haotian Lu +3
The multi-scale receptive field and large kernel attention (LKA) module have been shown to significantly improve performance in the lightweight image super-resolution task. However…