462 citations · 597 across the 9 of their papers we have counts for
13 papers
Cross Modal Compression: Towards Human-comprehensible Semantic Compression
Jiguo Li, Chuanmin Jia, Xinfeng Zhang +2
Traditional image/video compression aims to reduce the transmission/storage cost with signal fidelity as high as possible. However, with the increasing demand for machine analysis…
STAU: A SpatioTemporal-Aware Unit for Video Prediction and Beyond
Zheng Chang, Xinfeng Zhang, Shanshe Wang +2
Video prediction aims to predict future frames by modeling the complex spatiotemporal dynamics in videos. However, most of the existing methods only model the temporal information…
STRPM: A Spatiotemporal Residual Predictive Model for High-Resolution Video Prediction
Zheng Chang, Xinfeng Zhang, Shanshe Wang +2
Although many video prediction methods have obtained good performance in low-resolution (64128) videos, predictive models for high-resolution (5124K) videos have not be…
Improving Robustness and Accuracy via Relative Information Encoding in 3D Human Pose Estimation
Wenkang Shan, Haopeng Lu, Shanshe Wang +2
Most of the existing 3D human pose estimation approaches mainly focus on predicting 3D positional relationships between the root joint and other human joints (local motion) instead…
Region-adaptive Texture Enhancement for Detailed Person Image Synthesis
Lingbo Yang, Pan Wang, Xinfeng Zhang +6
The ability to produce convincing textural details is essential for the fidelity of synthesized person images. However, existing methods typically follow a ``warping-based'' strate…
User-generated Video Quality Assessment: A Subjective and Objective Study
Yang Li, Shengbin Meng, Xinfeng Zhang +3
Recently, we have observed an exponential increase of user-generated content (UGC) videos. The distinguished characteristic of UGC videos originates from the video production and d…