16 citations · 23 across the 2 of their papers we have counts for
5 papers
Detect an Object At Once without Fine-tuning
Junyu Hao, Jianheng Liu, Yongjia Zhao +5
When presented with one or a few photos of a previously unseen object, humans can instantly recognize it in different scenes. Although the human brain mechanism behind this phenome…
Robust Channel Learning for Large-Scale Radio Speaker Verification
Wenhao Yang, Jianguo Wei, Wenhuan Lu +2
Recent research in speaker verification has increasingly focused on achieving robust and reliable recognition under challenging channel conditions and noisy environments. Identifyi…
Attention Guidance Mechanism for Handwritten Mathematical Expression Recognition
Yutian Liu, Wenjun Ke, Jianguo Wei
Handwritten mathematical expression recognition (HMER) is challenging in image-to-text tasks due to the complex layouts of mathematical expressions and suffers from problems includ…
Constructing Holistic Spatio-Temporal Scene Graph for Video Semantic Role Labeling
Yu Zhao, Hao Fei, Yixin Cao +5
Video Semantic Role Labeling (VidSRL) aims to detect the salient events from given videos, by recognizing the predict-argument event structures and the interrelationships between e…
Generating Visual Spatial Description via Holistic 3D Scene Understanding
Yu Zhao, Hao Fei, Wei Ji +4
Visual spatial description (VSD) aims to generate texts that describe the spatial relations of the given objects within images. Existing VSD work merely models the 2D geometrical v…