12 papers
Robust Deepfake Detection, NTIRE 2026 Challenge: Report
Benedikt Hopf, Radu Timofte, Chenfan Qu +54
Robustness is a long-overlooked problem in deepfake detection. However, detection performance is nearly worthless in the real world if it suffers under exposure to even slight imag…
NTIRE 2026 Challenge on Single Image Reflection Removal in the Wild: Datasets, Results, and Methods
Jie Cai, Kangning Yang, Zhiyuan Li +50
In this paper, we review the NTIRE 2026 challenge on single-image reflection removal (SIRR) in the wild. SIRR is a fundamental task in image restoration. Despite progress in academ…
CANDLE: Illumination-Invariant Semantic Priors for Color Ambient Lighting Normalization
Rong-Lin Jian, Ting-Yao Chen, Yu-Fan Lin +4
Color ambient lighting normalization under multi-colored illumination is challenging due to severe chromatic shifts, highlight saturation, and material-dependent reflectance. Exist…
Frequency Switching Mechanism for Parameter-E!cient Multi-Task Learning
Shih-Wen Liu, Yen-Chang Chen, Wei-Ta Chu +2
Multi-task learning (MTL) aims to enable a single model to solve multiple tasks efficiently; however, current parameter-efficient fine-tuning (PEFT) methods remain largely limited…
VISTA: Validation-Guided Integration of Spatial and Temporal Foundation Models with Anatomical Decoding for Rare-Pathology VCE Event Detection
Bo-Cheng Qiu, Yu-Fan Lin, Yu-Zhe Pien +4
Capsule endoscopy event detection is challenging because diagnostically relevant findings are sparse, visually heterogeneous, and embedded in long, noisy video streams, while evalu…
Fast-ThinkAct: Efficient Vision-Language-Action Reasoning via Verbalizable Latent Planning
Chi-Pin Huang, Yunze Man, Zhiding Yu +4
Vision-Language-Action (VLA) tasks require reasoning over complex visual scenes and executing adaptive actions in dynamic environments. While recent studies on reasoning VLAs show…