activity
20242026
collaborators

9 papers

eess.SP2026

Object-Attribute-Relation Model Driven Adaptive Hierarchical Transmission for Multimodal Semantic Communication

Chenxing Li, Yiping Duan, Han Jiao +3

Traditional video coding (VVC, HEVC) prioritizes human visual perception, transmitting substantial texture redundancy that severely hinders machine decision-making under constraine…

eess.IV2026

Reinforced Rate Control for Neural Video Compression via Inter-Frame Rate-Distortion Awareness

Wuyang Cong, Junqi Shi, Lizhong Wang +4

Neural video compression (NVC) has demonstrated superior compression efficiency, yet effective rate control remains a significant challenge due to complex temporal dependencies. Ex…

cs.CV2026

ResTok: Learning Hierarchical Residuals in 1D Visual Tokenizers for Autoregressive Image Generation

Xu Zhang, Cheng Da, Huan Yang +3

Existing 1D visual tokenizers for autoregressive (AR) generation largely follow the design principles of language modeling, as they are built directly upon transformers whose prior…

eess.IV2026

YODA: Yet Another One-step Diffusion-based Video Compressor

Xingchen Li, Junzhe Zhang, Junqi Shi +2

While one-step diffusion models have recently excelled in perceptual image compression, their application to video remains limited. Prior efforts typically rely on pretrained 2D au…

cs.CV2025

Perception-Oriented Latent Coding for High-Performance Compressed Domain Semantic Inference

Xu Zhang, Ming Lu, Yan Chen +1

In recent years, compressed domain semantic inference has primarily relied on learned image coding models optimized for mean squared error (MSE). However, MSE-oriented optimization…

cs.MM2025

Adaptive Rate Control for Deep Video Compression with Rate-Distortion Prediction

Bowen Gu, Hao Chen, Ming Lu +2

Deep video compression has made significant progress in recent years, achieving rate-distortion performance that surpasses that of traditional video compression methods. However, r…