6 papers
NanoHTNet: Nano Human Topology Network for Efficient 3D Human Pose Estimation
Jialun Cai, Mengyuan Liu, Hong Liu +2
The widespread application of 3D human pose estimation (HPE) is limited by resource-constrained edge devices, requiring more efficient models. A key approach to enhancing efficienc…
UST-SSM: Unified Spatio-Temporal State Space Models for Point Cloud Video Modeling
Peiming Li, Ziyi Wang, Yulin Yuan +4
Point cloud videos capture dynamic 3D motion while reducing the effects of lighting and viewpoint variations, making them highly effective for recognizing subtle and continuous hum…
TCPFormer: Learning Temporal Correlation with Implicit Pose Proxy for 3D Human Pose Estimation
Jiajie Liu, Mengyuan Liu, Hong Liu +1
Recent multi-frame lifting methods have dominated the 3D human pose estimation. However, previous methods ignore the intricate dependence within the 2D pose sequence and learn sing…
Learning Mutual Excitation for Hand-to-Hand and Human-to-Human Interaction Recognition
Mengyuan Liu, Chen Chen, Songtao Wu +2
Recognizing interactive actions, including hand-to-hand interaction and human-to-human interaction, has attracted increasing attention for various applications in the field of vide…
Cross-Model Cross-Stream Learning for Self-Supervised Human Action Recognition
Mengyuan Liu, Hong Liu, Tianyu Guo
Considering the instance-level discriminative ability, contrastive learning methods, including MoCo and SimCLR, have been adapted from the original image representation learning ta…
ClickDiff: Click to Induce Semantic Contact Map for Controllable Grasp Generation with Diffusion Models
Peiming Li, Ziyi Wang, Mengyuan Liu +2
Grasp generation aims to create complex hand-object interactions with a specified object. While traditional approaches for hand generation have primarily focused on visibility and…