3 papers
cs.CV2024
High Efficiency Image Compression for Large Visual-Language Models
Binzhe Li, Shurun Wang, Shiqi Wang +1
In recent years, large visual language models (LVLMs) have shown impressive performance and promising generalization capability in multi-modal tasks, thus replacing humans as recei…
cs.CV2023
Semantic Face Compression for Metaverse: A Compact 3D Descriptor Based Approach
Binzhe Li, Bolin Chen, Zhao Wang +2
In this letter, we envision a new metaverse communication paradigm for virtual avatar faces, and develop the semantic face compression with compact 3D facial descriptors. The funda…
cs.CV2023
Interactive Face Video Coding: A Generative Compression Framework
Bolin Chen, Zhao Wang, Binzhe Li +3
In this paper, we propose a novel framework for Interactive Face Video Coding (IFVC), which allows humans to interact with the intrinsic visual representations instead of the signa…