42 citations · 64 across the 11 of their papers we have counts for
12 papers · 1 filter
Sub-action Prototype Learning for Point-level Weakly-supervised Temporal Action Localization
Yueyang Li, Yonghong Hou, Wanqing Li
Point-level weakly-supervised temporal action localization (PWTAL) aims to localize actions with only a single timestamp annotation for each action instance. Existing methods tend…
Focal and Global Spatial-Temporal Transformer for Skeleton-based Action Recognition
Zhimin Gao, Peitao Wang, Pei Lv +5
Despite great progress achieved by transformer in various vision tasks, it is still underexplored for skeleton-based action recognition with only a few attempts. Besides, these met…
FT-HID: A Large Scale RGB-D Dataset for First and Third Person Human Interaction Analysis
Zihui Guo, Yonghong Hou, Pichao Wang +3
Analysis of human interaction is one important research topic of human motion analysis. It has been studied either using first person vision (FPV) or third person vision (TPV). How…
Novel View Synthesis from a Single Image via Unsupervised learning
Bingzheng Liu, Jianjun Lei, Bo Peng +3
View synthesis aims to generate novel views from one or more given source views. Although existing methods have achieved promising performance, they usually require paired views of…
A Two-stream Neural Network for Pose-based Hand Gesture Recognition
Chuankun Li, Shuai Li, Yanbo Gao +2
Pose based hand gesture recognition has been widely studied in the recent years. Compared with full body action recognition, hand gesture involves joints that are more spatially cl…
Trear: Transformer-based RGB-D Egocentric Action Recognition
Xiangyu Li, Yonghong Hou, Pichao Wang +3
In this paper, we propose a \textbf{Tr}ansformer-based RGB-D \textbf{e}gocentric \textbf{a}ction \textbf{r}ecognition framework, called Trear. It consists of two modules, inter-fra…