activity
20202022
most citedComplex Sequential Understanding through the Awareness of Spatial and Temporal Concepts

27 citations · 73 across the 12 of their papers we have counts for

collaborators
Showing cs.CVShow all

10 papers · 1 filter

cs.CV2022

Semantic Segmentation by Early Region Proxy

Yifan Zhang, Bo Pang, Cewu Lu

Typical vision backbones manipulate structured features. As a compromise, semantic segmentation has long been modeled as per-point prediction on dense regular grids. In this work,…

cs.CV20211 cited

PGT: A Progressive Method for Training Models on Long Videos

Bo Pang, Gao Peng, Yizhuo Li +1

Convolutional video models have an order of magnitude larger computational complexity than their counterpart image-level models. Constrained by computational resources, there is no…

cs.CV20202 cited

TDAF: Top-Down Attention Framework for Vision Tasks

Bo Pang, Yizhuo Li, Jiefeng Li +3

Human attention mechanisms often work in a top-down manner, yet it is not well explored in vision research. Here, we propose the Top-Down Attention Framework (TDAF) to capture top-…

cs.CV202014 cited

Multimodal Pretraining for Dense Video Captioning

Gabriel Huang, Bo Pang, Zhenhai Zhu +2

Learning specific hands-on skills such as cooking, car maintenance, and home repairs increasingly happens via instructional videos. The user experience with such videos is known to…

cs.CV2020

Fully Unsupervised Person Re-identification viaSelective Contrastive Learning

Bo Pang, Deming Zhai, Junjun Jiang +1

Person re-identification (ReID) aims at searching the same identity person among images captured by various cameras. Unsupervised person ReID attracts a lot of attention recently,…

cs.CV20208 cited

ASAP-Net: Attention and Structure Aware Point Cloud Sequence Segmentation

Hanwen Cao, Yongyi Lu, Cewu Lu +3

Recent works of point clouds show that mulit-frame spatio-temporal modeling outperforms single-frame versions by utilizing cross-frame information. In this paper, we further improv…