54 citations · 57 across the 4 of their papers we have counts for
11 papers
Audio-Language Models for Audio-Centric Tasks: A Systematic Survey
Yi Su, Jisheng Bai, Qisheng Xu +2
Audio-Language Models (ALMs), trained on paired audio-text data, are designed to process, understand, and reason about audio-centric multimodal content. Unlike traditional supervis…
NTIRE 2025 Challenge on Image Super-Resolution (x4): Methods and Results
Zheng Chen, Kai Liu, Jue Gong +108
This paper presents the NTIRE 2025 image super-resolution (4) challenge, one of the associated competitions of the 10th NTIRE Workshop at CVPR 2025. The challenge aims to r…
NTIRE 2024 Challenge on Image Super-Resolution (x4): Methods and Results
Zheng Chen, Zongwei Wu, Eduard Zamfir +85
This paper reviews the NTIRE 2024 challenge on image super-resolution (4), highlighting the solutions proposed and the outcomes obtained. The challenge involves generating…
AudioDER: A Deduplication-Enhanced Reasoning Dataset for Post-Training Large Audio-Language Models
Hui Geng, Yi Su, Han Yin +7
Recent advances in pretrained large audio-language models (LALMs) have demonstrated strong capabilities across speech, sound, and music. To adapt these models to downstream tasks w…
NTIRE 2026 The Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images: Methods and Results
Xin Li, Yeying Jin, Suhang Yao +95
This paper presents an overview of the NTIRE 2026 Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images. Building upon the success of the first edition, this c…
The First Challenge on Remote Sensing Infrared Image Super-Resolution at NTIRE 2026: Benchmark Results and Method Overview
Kai Liu, Haoyang Yue, Zeli Lin +65
This paper presents the NTIRE 2026 Remote Sensing Infrared Image Super-Resolution (x4) Challenge, one of the associated challenges of NTIRE 2026. The challenge aims to recover high…