5 papers
A Lightweight Dual-Mode Optimization for Generative Face Video Coding
Zihan Zhang, Shanzhi Yin, Bolin Chen +3
Generative Face Video Coding (GFVC) achieves superior rate-distortion performance by leveraging the strong inference capabilities of deep generative models. However, its practical…
Rethinking Generative Human Video Coding with Implicit Motion Transformation
Bolin Chen, Ru-Ling Liao, Jie Chen +1
Beyond traditional hybrid-based video codec, generative video codec could achieve promising compression performance by evolving high-dimensional signals into compact feature repres…
Generative Models at the Frontier of Compression: A Survey on Generative Face Video Coding
Bolin Chen, Shanzhi Yin, Goluck Konuko +4
The rise of deep generative models has greatly advanced video compression, reshaping the paradigm of face video coding through their powerful capability for semantic-aware represen…
Latent Guidance in Diffusion Models for Perceptual Evaluations
Shreshth Saini, Ru-Ling Liao, Yan Ye +1
Despite recent advancements in latent diffusion models that generate high-dimensional image data and perform various downstream tasks, there has been little exploration into percep…
Pleno-Generation: A Scalable Generative Face Video Compression Framework with Bandwidth Intelligence
Bolin Chen, Hanwei Zhu, Shanzhi Yin +5
Generative model based compact video compression is typically operated within a relative narrow range of bitrates, and often with an emphasis on ultra-low rate applications. There…