96 citations · 679 across the 47 of their papers we have counts for
10 papers · 1 filter
Dynamic Fusion with Intra- and Inter- Modality Attention Flow for Visual Question Answering
Gao Peng, Zhengkai Jiang, Haoxuan You +4
Learning effective fusion of multi-modality features is at the heart of visual question answering. We propose a novel method of dynamically fusing multi-modal features with intra-…
Learning Monocular Depth by Distilling Cross-domain Stereo Networks
Xiaoyang Guo, Hongsheng Li, Shuai Yi +2
Monocular depth estimation aims at estimating a pixelwise depth map for a single image, which has wide applications in scene understanding and autonomous driving. Existing supervis…
HMS-Net: Hierarchical Multi-scale Sparsity-invariant Network for Sparse Depth Completion
Zixuan Huang, Junming Fan, Shenggan Cheng +3
Dense depth cues are important and have wide applications in various computer vision tasks. In autonomous driving, LIDAR sensors are adopted to acquire depth measurements around th…
Generative Adversarial Frontal View to Bird View Synthesis
Xinge Zhu, Zhichao Yin, Jianping Shi +2
Environment perception is an important task with great practical value and bird view is an essential part for creating panoramas of surrounding environment. Due to the large gap an…
Person Re-identification with Deep Similarity-Guided Graph Neural Network
Yantao Shen, Hongsheng Li, Shuai Yi +2
The person re-identification task requires to robustly estimate visual similarities between person images. However, existing person re-identification models mostly estimate the sim…
Deep Continuous Conditional Random Fields with Asymmetric Inter-object Constraints for Online Multi-object Tracking
Hui Zhou, Wanli Ouyang, Jian Cheng +2
Online Multi-Object Tracking (MOT) is a challenging problem and has many important applications including intelligence surveillance, robot navigation and autonomous driving. In exi…