activity
20192023
most citedHypergraph Transformer for Skeleton-based Action Recognition

31 citations · 160 across the 35 of their papers we have counts for

collaborators
Showing cs.CVShow all

43 papers · 1 filter

cs.CV2023★ 4 cited

Refined Temporal Pyramidal Compression-and-Amplification Transformer for 3D Human Pose Estimation

Hanbing Liu, Wangmeng Xiang, Jun-Yan He +4

Accurately estimating the 3D pose of humans in video sequences requires both accuracy and a well-structured architecture. With the success of transformers, we introduce the Refined…

cs.CV2023★ 1 cited

Towards Deeply Unified Depth-aware Panoptic Segmentation with Bi-directional Guidance Learning

Junwen He, Yifan Wang, Lijun Wang +6

Depth-aware panoptic segmentation is an emerging topic in computer vision which combines semantic and geometric understanding for more robust scene interpretation. Recent works pur…

cs.CV2023★ 11 cited

Pixel-Aware Stable Diffusion for Realistic Image Super-resolution and Personalized Stylization

Tao Yang, Rongyuan Wu, Peiran Ren +2

Diffusion models have demonstrated impressive performance in various image generation, editing, enhancement and translation tasks. In particular, the pre-trained text-to-image stab…

cs.CV2023★ 2 cited

TransFace++: Rethinking the Face Recognition Paradigm with a Focus on Accuracy, Efficiency, and Security

Jun Dan, Yang Liu, Baigui Sun +2

Face Recognition (FR) technology has made significant strides with the emergence of deep learning. Typically, most existing FR models are built upon Convolutional Neural Networks (…

cs.CV2023★ 2 cited

PoSynDA: Multi-Hypothesis Pose Synthesis Domain Adaptation for Robust 3D Human Pose Estimation

Hanbing Liu, Jun-Yan He, Zhi-Qi Cheng +8

Existing 3D human pose estimators face challenges in adapting to new datasets due to the lack of 2D-3D pose pairs in training sets. To overcome this issue, we propose \textit{Multi…

cs.CV2023★ 1 cited

CostFormer:Cost Transformer for Cost Aggregation in Multi-view Stereo

Weitao Chen, Hongbin Xu, Zhipeng Zhou +4

The core of Multi-view Stereo(MVS) is the matching process among reference and source pixels. Cost aggregation plays a significant role in this process, while previous methods focu…