6 papers
Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards
Yu Huang, Zihua Zhao, Zhaoxin Huan +9
The open-ended generation in LLMs usually requires multi-dimensional rubrics to adequately assess quality and guide the improvement of reinforcement learning. However, a critical d…
One-Step Diffusion Transformer for Controllable Real-World Image Super-Resolution
Yushun Fang, Yuxiang Chen, Shibo Yin +5
Recent advances in diffusion-based real-world image super-resolution (Real-ISR) have demonstrated remarkable perceptual quality, yet the balance between fidelity and controllabilit…
MoA-VR: A Mixture-of-Agents System Towards All-in-One Video Restoration
Lu Liu, Chunlei Cai, Shaocheng Shen +9
Real-world videos often suffer from complex degradations, such as noise, compression artifacts, and low-light distortions, due to diverse acquisition and transmission conditions. E…
4DGCPro: Efficient Hierarchical 4D Gaussian Compression for Progressive Volumetric Video Streaming
Zihan Zheng, Zhenlong Wu, Houqiang Zhong +7
Achieving seamless viewing of high-fidelity volumetric video, comparable to 2D video experiences, remains an open challenge. Existing volumetric video compression methods either la…
Instance-aware Image Colorization with Controllable Textual Descriptions and Segmentation Masks
Yanru An, Ling Gui, Chunlei Cai +5
Recently, the application of deep learning in image colorization has received widespread attention. The maturation of diffusion models has further advanced the development of image…
Serial Low-rank Adaptation of Vision Transformer
Houqiang Zhong, Shaocheng Shen, Ke Cai +7
Fine-tuning large pre-trained vision foundation models in a parameter-efficient manner is critical for downstream vision tasks, considering the practical constraints of computation…