3 citations · 3 across the 2 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Language-driven Description Generation and Common Sense Reasoning for Video Action Recognition
Xiaodan Hu, Chuhang Zou, Suchen Wang +2
Recent video action recognition methods have shown excellent performance by adapting large-scale pre-trained language-image models to the video domain. However, language models con…
cs.CV2021★ 3 cited
Unsupervised 3D Pose Estimation for Hierarchical Dance Video Recognition
Xiaodan Hu, Narendra Ahuja
Dance experts often view dance as a hierarchy of information, spanning low-level (raw images, image sequences), mid-levels (human poses and bodypart movements), and high-level (dan…