27 citations · 91 across the 14 of their papers we have counts for
11 papers · 1 filter
TDAF: Top-Down Attention Framework for Vision Tasks
Bo Pang, Yizhuo Li, Jiefeng Li +3
Human attention mechanisms often work in a top-down manner, yet it is not well explored in vision research. Here, we propose the Top-Down Attention Framework (TDAF) to capture top-…
Multimodal Pretraining for Dense Video Captioning
Gabriel Huang, Bo Pang, Zhenhai Zhu +2
Learning specific hands-on skills such as cooking, car maintenance, and home repairs increasingly happens via instructional videos. The user experience with such videos is known to…
Fully Unsupervised Person Re-identification viaSelective Contrastive Learning
Bo Pang, Deming Zhai, Junjun Jiang +1
Person re-identification (ReID) aims at searching the same identity person among images captured by various cameras. Unsupervised person ReID attracts a lot of attention recently,…
ASAP-Net: Attention and Structure Aware Point Cloud Sequence Segmentation
Hanwen Cao, Yongyi Lu, Cewu Lu +3
Recent works of point clouds show that mulit-frame spatio-temporal modeling outperforms single-frame versions by utilizing cross-frame information. In this paper, we further improv…
Robust Reinforcement Learning: A Case Study in Linear Quadratic Regulation
Bo Pang, Zhong-Ping Jiang
This paper studies the robustness of reinforcement learning algorithms to errors in the learning process. Specifically, we revisit the benchmark problem of discrete-time linear qua…
NTIRE 2020 Challenge on Video Quality Mapping: Methods and Results
Dario Fuoli, Zhiwu Huang, Martin Danelljan +18
This paper reviews the NTIRE 2020 challenge on video quality mapping (VQM), which addresses the issues of quality mapping from source video domain to target video domain. The chall…