50 citations · 66 across the 5 of their papers we have counts for
13 papers
Video-ReTime: Learning Temporally Varying Speediness for Time Remapping
Simon Jenni, Markus Woodson, Fabian Caba Heilbron
We propose a method for generating a temporally remapped video that matches the desired target duration while maximally preserving natural video dynamics. Our approach trains a neu…
Learning to Cut by Watching Movies
Alejandro Pardo, Fabian Caba Heilbron, Juan León Alcázar +2
Video content creation keeps growing at an incredible pace; yet, creating engaging stories remains challenging and requires non-trivial video editing expertise. Many video editing…
APES: Audiovisual Person Search in Untrimmed Video
Juan Leon Alcazar, Long Mai, Federico Perazzi +4
Humans are arguably one of the most important subjects in video streams, many real-world applications such as video summarization or video editing workflows often require the autom…
MAAS: Multi-modal Assignation for Active Speaker Detection
Juan León-Alcázar, Fabian Caba Heilbron, Ali Thabet +1
Active speaker detection requires a solid integration of multi-modal cues. While individual modalities can approximate a solution, accurate predictions can only be achieved by expl…
Real-time Semantic Segmentation with Fast Attention
Ping Hu, Federico Perazzi, Fabian Caba Heilbron +4
In deep CNN based models for semantic segmentation, high accuracy relies on rich spatial context (large receptive fields) and fine spatial details (high resolution), both of which…
Active Speakers in Context
Juan Leon Alcazar, Fabian Caba Heilbron, Long Mai +4
Current methods for active speak er detection focus on modeling short-term audiovisual information from a single speaker. Although this strategy can be enough for addressing single…