collaborators

6 papers

cs.CV2026

A Frozen Pixel-Space Diffusion Model Can Guide Itself with Its Own Samples

Zixuan Fu, Chong Wang, Lanqing Guo +3

Pixel-space diffusion models aim to learn an end-to-end generator directly over raw pixels. This is challenging because a single model must capture both global structure and local…

cs.CV2026

Frames2Residual: Spatiotemporal Decoupling for Self-Supervised Video Denoising

Mingjie Ji, Zhan Shi, Kailai Zhou +2

Self-supervised video denoising methods typically extend image-based frameworks into the temporal dimension, yet they often struggle to integrate inter-frame temporal consistency w…

cs.CV2026

ConceptSeg-R1: Segment Any Concept via Meta-Reinforcement Learning

Yuan Zhao, Youwei Pang, Jiaming Zuo +10

Recent progress in promptable segmentation has shifted visual perception from object-level localization toward concept-level understanding. However, the notion of a concept remains…

cs.CV2025

M-SpecGene: Generalized Foundation Model for RGBT Multispectral Vision

Kailai Zhou, Fuqiang Yang, Shixian Wang +5

RGB-Thermal (RGBT) multispectral vision is essential for robust perception in complex environments. Most RGBT tasks follow a case-by-case research paradigm, relying on manually cus…

cs.CV2025

Gaseous Object Detection

Kailai Zhou, Yibo Wang, Tao Lv +2

Object detection, a fundamental and challenging problem in computer vision, has experienced rapid development due to the effectiveness of deep learning. The current objects to be d…

cs.CV2024

Joint RGB-Spectral Decomposition Model Guided Image Enhancement in Mobile Photography

Kailai Zhou, Lijing Cai, Yibo Wang +4

The integration of miniaturized spectrometers into mobile devices offers new avenues for image quality enhancement and facilitates novel downstream tasks. However, the broader appl…