6 papers
OctoT2I: A Self-Evolving Agentic Text-to-Image Router
Xu Jiang, Bin Chen, Gehui Li +3
The explosive growth of Text-to-Image (T2I) models, from large-scale versions to lightweight, real-time ones, now faces diminishing marginal returns from single-model scaling. Agen…
VQ-Jarvis: Retrieval-Augmented Video Restoration Agent with Sharp Vision and Fast Thought
Xuanyu Zhang, Weiqi Li, Qunliang Xing +6
Video restoration in real-world scenarios is challenged by heterogeneous degradations, where static architectures and fixed inference pipelines often fail to generalize. Recent age…
Improved Adversarial Diffusion Compression for Real-World Video Super-Resolution
Bin Chen, Weiqi Li, Shijie Zhao +4
While many diffusion models have achieved impressive results in real-world video super-resolution (Real-VSR) by generating rich and realistic details, their reliance on multi-step…
LAFR: Efficient Diffusion-based Blind Face Restoration via Latent Codebook Alignment Adapter
Runyi Li, Bin Chen, Jian Zhang +1
Blind face restoration from low-quality (LQ) images is a challenging task that requires not only high-fidelity image reconstruction but also the preservation of facial identity. Wh…
CTSR: Controllable Fidelity-Realness Trade-off Distillation for Real-World Image Super Resolution
Runyi Li, Bin Chen, Jian Zhang +1
Real-world image super-resolution is a critical image processing task, where two key evaluation criteria are the fidelity to the original image and the visual realness of the gener…
Multi-Agent Image Restoration
Xu Jiang, Gehui Li, Bin Chen +1
Image restoration (IR) is challenging due to the complexity of real-world degradations. While many specialized and all-in-one IR models have been developed, they fail to effectivel…