5 papers
URVC: A Unified Real-Time Neural Video Coding Model with Temporal, Spatial, and Perceptual Adaptivity
Xihua Sheng, Chang Wen Chen
The paper introduces URVC, a real‑time neural video codec that can adapt its temporal prediction, spatial bit allocation, and perceptual quality on the fly, eliminating the need fo…
CADC: Content Adaptive Diffusion-Based Generative Image Compression
Xihua Sheng, Lingyu Zhu, Tianyu Zhang +3
Diffusion-based generative image compression has demonstrated remarkable potential for achieving realistic reconstruction at ultra-low bitrates. The key to unlocking this potential…
Fine-Grained Motion Compression and Selective Temporal Fusion for Neural B-Frame Video Coding
Xihua Sheng, Peilin Chen, Meng Wang +3
With the remarkable progress in neural P-frame video coding, neural B-frame coding has recently emerged as a critical research direction. However, most existing neural B-frame code…
DCVC-MV: Deep Contextual Multiview Video Compression with Efficient Inter-View Prediction
Xihua Sheng, Yingwen Zhang, Long Xu +1
Multiview video is a key format for 3D applications such as free-viewpoint broadcasting and virtual reality, yet its large data volume poses significant challenges for efficient st…
An Information-Theoretic Regularizer for Lossy Neural Image Compression
Yingwen Zhang, Meng Wang, Xihua Sheng +4
Lossy image compression networks aim to minimize the latent entropy of images while adhering to specific distortion constraints. However, optimizing the neural network can be chall…