5 papers
A Lightweight Dual-Mode Optimization for Generative Face Video Coding
Zihan Zhang, Shanzhi Yin, Bolin Chen +3
Generative Face Video Coding (GFVC) achieves superior rate-distortion performance by leveraging the strong inference capabilities of deep generative models. However, its practical…
Generative Models at the Frontier of Compression: A Survey on Generative Face Video Coding
Bolin Chen, Shanzhi Yin, Goluck Konuko +4
The rise of deep generative models has greatly advanced video compression, reshaping the paradigm of face video coding through their powerful capability for semantic-aware represen…
Compressing Human Body Video with Interactive Semantics: A Generative Approach
Bolin Chen, Shanzhi Yin, Hanwei Zhu +6
In this paper, we propose to compress human body video with interactive semantics, which can facilitate video coding to be interactive and controllable by manipulating semantic-lev…
Beyond GFVC: A Progressive Face Video Compression Framework with Adaptive Visual Tokens
Bolin Chen, Shanzhi Yin, Zihan Zhang +5
Recently, deep generative models have greatly advanced the progress of face video coding towards promising rate-distortion performance and diverse application functionalities. Beyo…
Tokenizing Motion: A Generative Approach for Scene Dynamics Compression
Shanzhi Yin, Zihan Zhang, Bolin Chen +2
This paper proposes a novel generative video compression framework that leverages motion pattern priors, derived from subtle dynamics in common scenes (e.g., swaying flowers or a b…