Showing cs.CVShow all
3 papers · 1 filter
cs.CV2020
Active Speakers in Context
Juan Leon Alcazar, Fabian Caba Heilbron, Long Mai +4
Current methods for active speak er detection focus on modeling short-term audiovisual information from a single speaker. Although this strategy can be enough for addressing single…
cs.CV2018
SMIT: Stochastic Multi-Label Image-to-Image Translation
Andrés Romero, Pablo Arbeláez, Luc Van Gool +1
Cross-domain mapping has been a very active topic in recent years. Given one image, its main purpose is to translate it to the desired target domain, or multiple domains in the cas…
cs.CV2018
Dynamic Multimodal Instance Segmentation guided by natural language queries
Edgar Margffoy-Tuay, Juan C. Pérez, Emilio Botero +1
We address the problem of segmenting an object given a natural language expression that describes it. Current techniques tackle this task by either (\textit{i}) directly or recursi…