5 citations · 9 across the 7 of their papers we have counts for
7 papers
Image Compression for Machine and Human Vision with Spatial-Frequency Adaptation
Han Li, Shaohui Li, Shuangrui Ding +5
Image compression for machine and human vision (ICMH) has gained increasing attention in recent years. Existing ICMH methods are limited by high training and storage overheads due…
AiluRus: A Scalable ViT Framework for Dense Prediction
Jin Li, Yaoming Wang, Xiaopeng Zhang +6
Vision transformers (ViTs) have emerged as a prevalent architecture for vision tasks owing to their impressive performance. However, when it comes to handling long token sequences,…
ActionPrompt: Action-Guided 3D Human Pose Estimation With Text and Pose Prompting
Hongwei Zheng, Han Li, Bowen Shi +5
Recent 2D-to-3D human pose estimation (HPE) utilizes temporal consistency across sequences to alleviate the depth ambiguity problem but ignore the action related prior knowledge hi…
Scene Graph Lossless Compression with Adaptive Prediction for Objects and Relations
Yufeng Zhang, Weiyao Lin, Wenrui Dai +2
The scene graph is a new data structure describing objects and their pairwise relationship within image scenes. As the size of scene graph in vision applications grows, how to loss…
Learned Lossless Compression for JPEG via Frequency-Domain Prediction
Jixiang Luo, Shaohui Li, Wenrui Dai +3
JPEG images can be further compressed to enhance the storage and transmission of large-scale image datasets. Existing learned lossless compressors for RGB images cannot be well tra…
Pose-Oriented Transformer with Uncertainty-Guided Refinement for 2D-to-3D Human Pose Estimation
Han Li, Bowen Shi, Wenrui Dai +7
There has been a recent surge of interest in introducing transformers to 3D human pose estimation (HPE) due to their powerful capabilities in modeling long-term dependencies. Howev…