1 citations · 1 across the 3 of their papers we have counts for
3 papers · 1 filter
Transcription-Enriched Joint Embeddings for Spoken Descriptions of Images and Videos
Benet Oriol, Jordi Luque, Ferran Diego +1
In this work, we propose an effective approach for training unique embedding representations by combining three simultaneous modalities: image and spoken and textual narratives. Th…
Unsupervised Representation Learning by Discovering Reliable Image Relations
Timo Milbich, Omair Ghori, Ferran Diego +1
Learning robust representations that allow to reliably establish relations between images is of paramount importance for virtually all of computer vision. Annotating the quadratic…
CEREALS - Cost-Effective REgion-based Active Learning for Semantic Segmentation
Radek Mackowiak, Philip Lenz, Omair Ghori +3
State of the art methods for semantic image segmentation are trained in a supervised fashion using a large corpus of fully labeled training images. However, gathering such a corpus…