3 citations · 6 across the 4 of their papers we have counts for
9 papers
SBAT: Video Captioning with Sparse Boundary-Aware Transformer
Tao Jin, Siyu Huang, Ming Chen +2
In this paper, we focus on the problem of applying the transformer structure to video captioning effectively. The vanilla transformer is proposed for uni-modal language generation…
Multitask Non-Autoregressive Model for Human Motion Prediction
Bin Li, Jian Tian, Zhongfei Zhang +2
Human motion prediction, which aims at predicting future human skeletons given the past ones, is a typical sequence-to-sequence problem. Therefore, extensive efforts have been cont…
Perceiving Physical Equation by Observing Visual Scenarios
Siyu Huang, Zhi-Qi Cheng, Xi Li +3
Inferring universal laws of the environment is an important ability of human intelligence as well as a symbol of general AI. In this paper, we take a step toward this goal such tha…
Stacked Pooling: Improving Crowd Counting by Boosting Scale Invariance
Siyu Huang, Xi Li, Zhi-Qi Cheng +2
In this work, we explore the cross-scale similarity in crowd counting scenario, in which the regions of different scales often exhibit high visual similarity. This feature is unive…
Multi-Channel Pyramid Person Matching Network for Person Re-Identification
Chaojie Mao, Yingming Li, Yaqing Zhang +2
In this work, we present a Multi-Channel deep convolutional Pyramid Person Matching Network (MC-PPMN) based on the combination of the semantic-components and the color-texture dist…
Boosted Zero-Shot Learning with Semantic Correlation Regularization
Te Pi, Xi Li, Zhongfei +1
We study zero-shot learning (ZSL) as a transfer learning problem, and focus on the two key aspects of ZSL, model effectiveness and model adaptation. For effective modeling, we adop…