72 citations · 115 across the 7 of their papers we have counts for
14 papers · 1 filter
Compositional Sketch Search
Alexander Black, Tu Bui, Long Mai +2
We present an algorithm for searching image collections using free-hand sketches that describe the appearance and relative positions of multiple objects. Sketch based image retriev…
APES: Audiovisual Person Search in Untrimmed Video
Juan Leon Alcazar, Long Mai, Federico Perazzi +4
Humans are arguably one of the most important subjects in video streams, many real-world applications such as video summarization or video editing workflows often require the autom…
Boosting Monocular Depth Estimation Models to High-Resolution via Content-Adaptive Multi-Resolution Merging
S. Mahdi H. Miangoleh, Sebastian Dille, Long Mai +2
Neural networks have shown great abilities in estimating depth from a single image. However, the inferred depth maps are well below one-megapixel resolution and often lack fine-gra…
Learning to Recover 3D Scene Shape from a Single Image
Wei Yin, Jianming Zhang, Oliver Wang +4
Despite significant progress in monocular depth estimation in the wild, recent state-of-the-art methods cannot be used to recover accurate 3D scene shape due to an unknown depth sh…
Revisiting Adaptive Convolutions for Video Frame Interpolation
Simon Niklaus, Long Mai, Oliver Wang
Video frame interpolation, the synthesis of novel views in time, is an increasingly popular research direction with many new papers further advancing the state of the art. But as e…
Active Speakers in Context
Juan Leon Alcazar, Fabian Caba Heilbron, Long Mai +4
Current methods for active speak er detection focus on modeling short-term audiovisual information from a single speaker. Although this strategy can be enough for addressing single…