2 citations · 4 across the 4 of their papers we have counts for
4 papers
VisAug: Facilitating Speech-Rich Web Video Navigation and Engagement with Auto-Generated Visual Augmentations
Baoquan Zhao, Xiaofan Ma, Qianshi Pang +3
The widespread adoption of digital technology has ushered in a new era of digital transformation across all aspects of our lives. Online learning, social, and work activities, such…
DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition
Xinzhu Li, Juepeng Zheng, Yikun Chen +7
Robust gait recognition requires highly discriminative representations, which are closely tied to input modalities. While binary silhouettes and skeletons have dominated recent lit…
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering
Yiran Meng, Junhong Ye, Wei Zhou +4
Cross-video question answering presents significant challenges beyond traditional single-video understanding, particularly in establishing meaningful connections across video strea…
MorphText: Deep Morphology Regularized Arbitrary-shape Scene Text Detection
Chengpei Xu, Wenjing Jia, Ruomei Wang +2
Bottom-up text detection methods play an important role in arbitrary-shape scene text detection but there are two restrictions preventing them from achieving their great potential,…