1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2026
SBF: An Effective Representation to Augment Skeleton for Video-based Human Action Recognition
Zhuoxuan Peng, Yiyi Ding, Yang Lin +1
Many modern video-based human action recognition (HAR) approaches use 2D skeleton as the intermediate representation in their prediction pipelines. Despite overall encouraging resu…
cs.CV2026
Expanding mmWave Datasets for Human Pose Estimation with Unlabeled Data and LiDAR Datasets
Zhuoxuan Peng, Boan Zhu, Xingjian Zhang +2
Current millimeter-wave (mmWave) datasets for human pose estimation (HPE) are scarce and lack diversity in both point cloud (PC) attributes and human poses, hindering the generaliz…
cs.CV2024★ 1 cited
Revisiting Referring Expression Comprehension Evaluation in the Era of Large Multimodal Models
Jierun Chen, Fangyun Wei, Jinjing Zhao +5
Referring expression comprehension (REC) involves localizing a target instance based on a textual description. Recent advancements in REC have been driven by large multimodal model…