3 citations · 7 across the 5 of their papers we have counts for
5 papers
Learning Procedure-aware Video Representation from Instructional Videos and Their Narrations
Yiwu Zhong, Licheng Yu, Yang Bai +3
The abundance of instructional videos and their narrations over the Internet offers an exciting avenue for understanding procedural activities. In this work, we propose to learn vi…
Temporal Segment Transformer for Action Segmentation
Zhichao Liu, Leshan Wang, Desen Zhou +5
Recognizing human actions from untrimmed videos is an important task in activity understanding, and poses unique challenges in modeling long-range temporal relations. Recent works…
WhisperWand: Simultaneous Voice and Gesture Tracking Interface
Yang Bai, Irtaza Shahid, Harshvardhan Takawale +1
This paper presents the design and implementation of WhisperWand, a comprehensive voice and motion tracking interface for voice assistants. Distinct from prior works, WhisperWand i…
Action Quality Assessment with Temporal Parsing Transformer
Yang Bai, Desen Zhou, Songyang Zhang +5
Action Quality Assessment(AQA) is important for action understanding and resolving the task poses unique challenges due to subtle visual differences. Existing state-of-the-art meth…
Automated Customization of On-Thing Inference for Quality-of-Experience Enhancement
Yang Bai, Lixing Chen, Shaolei Ren +1
The rapid uptake of intelligent applications is pushing deep learning (DL) capabilities to Internet-of-Things (IoT). Despite the emergence of new tools for embedding deep neural ne…