87 citations · 111 across the 7 of their papers we have counts for
5 papers · 1 filter
Reasoning over Vision and Language: Exploring the Benefits of Supplemental Knowledge
Violetta Shevchenko, Damien Teney, Anthony Dick +1
The limits of applicability of vision-and-language models are defined by the coverage of their training data. Tasks like vision question answering (VQA) often require commonsense a…
Visual Question Answering with Prior Class Semantics
Violetta Shevchenko, Damien Teney, Anthony Dick +1
We present a novel mechanism to embed prior knowledge in a model for visual question answering. The open-set nature of the task is at odds with the ubiquitous approach of training…
Joint Learning of Set Cardinality and State Distribution
S. Hamid Rezatofighi, Anton Milan, Qinfeng Shi +2
We present a novel approach for learning to predict sets using deep learning. In recent years, deep neural networks have shown remarkable results in computer vision, natural langua…
A Survey of Appearance Models in Visual Object Tracking
Xi Li, Weiming Hu, Chunhua Shen +3
Visual object tracking is a significant computer vision task which can be applied to many domains such as visual surveillance, human computer interaction, and video compression. In…
Incremental Learning of 3D-DCT Compact Representations for Robust Visual Tracking
Xi Li, Anthony Dick, Chunhua Shen +2
Visual tracking usually requires an object appearance model that is robust to changing illumination, pose and other factors encountered in video. In this paper, we construct an app…