activity
20242026
collaborators

31 papers

cs.CV2026

Towards Compact Unified Multimodal Tracking: Synergizing Knowledge Distillation with Structural Pruning

Yuqi Li, Yuedong Tan, Huiran Duan +7

Unified multimodal object tracking has achieved remarkable robustness by leveraging complementary sensor data (e.g., RGB, Thermal, Depth), yet the heavy computational burden of sta…

cs.CV2026

TRaM-VSR: Importance-Aware Token Routing and Merging for One-Step Diffusion Video Super-Resolution

Sicheng Gao, Zhuyun Zhou, Yixuan Liu +3

Video super-resolution (VSR) using large-scale Diffusion Transformer (DiT) priors achieves exceptional perceptual quality but is often impractical due to the quadratic computationa…

cs.CV2026

NTIRE 2025 Challenge on Image Super-Resolution (x4): Methods and Results

Zheng Chen, Kai Liu, Jue Gong +108

This paper presents the NTIRE 2025 image super-resolution (4) challenge, one of the associated competitions of the 10th NTIRE Workshop at CVPR 2025. The challenge aims to r…

cs.CV2026

NTIRE 2024 Challenge on Image Super-Resolution (x4): Methods and Results

Zheng Chen, Zongwei Wu, Eduard Zamfir +85

This paper reviews the NTIRE 2024 challenge on image super-resolution (4), highlighting the solutions proposed and the outcomes obtained. The challenge involves generating…

cs.CV2026

The Regularizing Power of Language-Training Deepfake Detectors

Benedikt Hopf, Zongwei Wu, Radu Timofte

Recently, thanks to the advent of Multimodal-LLMs, deepfake detectors are striving not only to be generalizable but also interpretable. We propose that these two challenges can eff…

cs.CV2026

NTIRE 2026 The Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images: Methods and Results

Xin Li, Yeying Jin, Suhang Yao +95

This paper presents an overview of the NTIRE 2026 Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images. Building upon the success of the first edition, this c…