31 citations · 57 across the 7 of their papers we have counts for
5 papers
1st Place Solution for PVUW Challenge 2023: Video Panoptic Segmentation
Tao Zhang, Xingye Tian, Haoran Wei +6
Video panoptic segmentation is a challenging task that serves as the cornerstone of numerous downstream applications, including video editing and autonomous driving. We believe tha…
MobRecon: Mobile-Friendly Hand Mesh Reconstruction from Monocular Image
Xingyu Chen, Yufeng Liu, Yajiao Dong +5
In this work, we propose a framework for single-view hand mesh reconstruction, which can simultaneously achieve high reconstruction accuracy, fast inference speed, and temporal coh…
A free lunch from ViT:Adaptive Attention Multi-scale Fusion Transformer for Fine-grained Visual Recognition
Yuan Zhang, Jian Cao, Ling Zhang +4
Learning subtle representation about object parts plays a vital role in fine-grained visual recognition (FGVR) field. The vision transformer (ViT) achieves promising results on com…
AdaPruner: Adaptive Channel Pruning and Effective Weights Inheritance
Xiangcheng Liu, Jian Cao, Hongyi Yao +2
Channel pruning is one of the major compression approaches for deep neural networks. While previous pruning methods have mostly focused on identifying unimportant channels, channel…
Recent Standard Development Activities on Video Coding for Machines
Wen Gao, Shan Liu, Xiaozhong Xu +3
In recent years, video data has dominated internet traffic and becomes one of the major data formats. With the emerging 5G and internet of things (IoT) technologies, more and more…