4 papers
Can LLMs Design Video Coding Tools? A Case Study on Planar Mode
Yingwen Zhang, Meng Wang, Liqiang He +1
This paper explores whether large language models (LLMs) can design video coding tools, a highly challenging task due to the intricate algorithmic coupling of tool modifications. I…
LLM-Driven Heuristic Frame-Level Quantization Parameter Adaptation for VVenC
Liqiang He, Yingwen Zhang, Riyu Lu +2
Optimal frame-level quantization parameter (QP) allocation remains a persistent challenge in modern video encoders. The fixed-QP scheme widely adopted in practical systems is inher…
Compact Visual Data Representation for Green Multimedia -- A Human Visual System Perspective
Peilin Chen, Xiaohan Fang, Meng Wang +2
The Human Visual System (HVS), with its intricate sophistication, is capable of achieving ultra-compact information compression for visual signals. This remarkable ability is coupl…
Scalable Face Image Coding via StyleGAN Prior: Towards Compression for Human-Machine Collaborative Vision
Qi Mao, Chongyu Wang, Meng Wang +4
The accelerated proliferation of visual content and the rapid development of machine vision technologies bring significant challenges in delivering visual data on a gigantic scale,…