activity
20172021
most citedVideo Frame Interpolation via Adaptive Separable Convolution

72 citations · 115 across the 7 of their papers we have counts for

collaborators
Showing cs.CVShow all

14 papers · 1 filter

cs.CV2021

Compositional Sketch Search

Alexander Black, Tu Bui, Long Mai +2

We present an algorithm for searching image collections using free-hand sketches that describe the appearance and relative positions of multiple objects. Sketch based image retriev…

cs.CV2021

APES: Audiovisual Person Search in Untrimmed Video

Juan Leon Alcazar, Long Mai, Federico Perazzi +4

Humans are arguably one of the most important subjects in video streams, many real-world applications such as video summarization or video editing workflows often require the autom…

cs.CV20211 cited

Boosting Monocular Depth Estimation Models to High-Resolution via Content-Adaptive Multi-Resolution Merging

S. Mahdi H. Miangoleh, Sebastian Dille, Long Mai +2

Neural networks have shown great abilities in estimating depth from a single image. However, the inferred depth maps are well below one-megapixel resolution and often lack fine-gra…

cs.CV20204 cited

Learning to Recover 3D Scene Shape from a Single Image

Wei Yin, Jianming Zhang, Oliver Wang +4

Despite significant progress in monocular depth estimation in the wild, recent state-of-the-art methods cannot be used to recover accurate 3D scene shape due to an unknown depth sh…

cs.CV20202 cited

Revisiting Adaptive Convolutions for Video Frame Interpolation

Simon Niklaus, Long Mai, Oliver Wang

Video frame interpolation, the synthesis of novel views in time, is an increasingly popular research direction with many new papers further advancing the state of the art. But as e…

cs.CV2020

Active Speakers in Context

Juan Leon Alcazar, Fabian Caba Heilbron, Long Mai +4

Current methods for active speak er detection focus on modeling short-term audiovisual information from a single speaker. Although this strategy can be enough for addressing single…