72 citations · 115 across the 7 of their papers we have counts for
16 papers
Compositional Sketch Search
Alexander Black, Tu Bui, Long Mai +2
We present an algorithm for searching image collections using free-hand sketches that describe the appearance and relative positions of multiple objects. Sketch based image retriev…
APES: Audiovisual Person Search in Untrimmed Video
Juan Leon Alcazar, Long Mai, Federico Perazzi +4
Humans are arguably one of the most important subjects in video streams, many real-world applications such as video summarization or video editing workflows often require the autom…
Boosting Monocular Depth Estimation Models to High-Resolution via Content-Adaptive Multi-Resolution Merging
S. Mahdi H. Miangoleh, Sebastian Dille, Long Mai +2
Neural networks have shown great abilities in estimating depth from a single image. However, the inferred depth maps are well below one-megapixel resolution and often lack fine-gra…
Learning to Recover 3D Scene Shape from a Single Image
Wei Yin, Jianming Zhang, Oliver Wang +4
Despite significant progress in monocular depth estimation in the wild, recent state-of-the-art methods cannot be used to recover accurate 3D scene shape due to an unknown depth sh…
Out of Order: How Important Is The Sequential Order of Words in a Sentence in Natural Language Understanding Tasks?
Thang M. Pham, Trung Bui, Long Mai +1
Do state-of-the-art natural language understanding models care about word order - one of the most important characteristics of a sequence? Not always! We found 75% to 90% of the co…
Revisiting Adaptive Convolutions for Video Frame Interpolation
Simon Niklaus, Long Mai, Oliver Wang
Video frame interpolation, the synthesis of novel views in time, is an increasingly popular research direction with many new papers further advancing the state of the art. But as e…