2 papers
cs.CV2022
OTPose: Occlusion-Aware Transformer for Pose Estimation in Sparsely-Labeled Videos
Kyung-Min Jin, Gun-Hee Lee, Seong-Whan Lee
Although many approaches for multi-human pose estimation in videos have shown profound results, they require densely annotated data which entails excessive man labor. Furthermore,…
cs.CV2022
HTNet: Anchor-free Temporal Action Localization with Hierarchical Transformers
Tae-Kyung Kang, Gun-Hee Lee, Seong-Whan Lee
Temporal action localization (TAL) is a task of identifying a set of actions in a video, which involves localizing the start and end frames and classifying each action instance. Ex…