activity
20132020
most citedSBAT: Video Captioning with Sparse Boundary-Aware Transformer

3 citations · 6 across the 4 of their papers we have counts for

collaborators

9 papers

cs.CV20203 cited

SBAT: Video Captioning with Sparse Boundary-Aware Transformer

Tao Jin, Siyu Huang, Ming Chen +2

In this paper, we focus on the problem of applying the transformer structure to video captioning effectively. The vanilla transformer is proposed for uni-modal language generation…

cs.CV2020

Multitask Non-Autoregressive Model for Human Motion Prediction

Bin Li, Jian Tian, Zhongfei Zhang +2

Human motion prediction, which aims at predicting future human skeletons given the past ones, is a typical sequence-to-sequence problem. Therefore, extensive efforts have been cont…

cs.AI2018

Perceiving Physical Equation by Observing Visual Scenarios

Siyu Huang, Zhi-Qi Cheng, Xi Li +3

Inferring universal laws of the environment is an important ability of human intelligence as well as a symbol of general AI. In this paper, we take a step toward this goal such tha…

cs.CV2018

Stacked Pooling: Improving Crowd Counting by Boosting Scale Invariance

Siyu Huang, Xi Li, Zhi-Qi Cheng +2

In this work, we explore the cross-scale similarity in crowd counting scenario, in which the regions of different scales often exhibit high visual similarity. This feature is unive…

cs.CV2018

Multi-Channel Pyramid Person Matching Network for Person Re-Identification

Chaojie Mao, Yingming Li, Yaqing Zhang +2

In this work, we present a Multi-Channel deep convolutional Pyramid Person Matching Network (MC-PPMN) based on the combination of the semantic-components and the color-texture dist…

cs.LG20172 cited

Boosted Zero-Shot Learning with Semantic Correlation Regularization

Te Pi, Xi Li, Zhongfei +1

We study zero-shot learning (ZSL) as a transfer learning problem, and focus on the two key aspects of ZSL, model effectiveness and model adaptation. For effective modeling, we adop…