activity
20172022
most citedActivityNet Challenge 2017 Summary

50 citations · 66 across the 5 of their papers we have counts for

collaborators

13 papers

cs.CV2022

Video-ReTime: Learning Temporally Varying Speediness for Time Remapping

Simon Jenni, Markus Woodson, Fabian Caba Heilbron

We propose a method for generating a temporally remapped video that matches the desired target duration while maximally preserving natural video dynamics. Our approach trains a neu…

cs.CV2021

Learning to Cut by Watching Movies

Alejandro Pardo, Fabian Caba Heilbron, Juan León Alcázar +2

Video content creation keeps growing at an incredible pace; yet, creating engaging stories remains challenging and requires non-trivial video editing expertise. Many video editing…

cs.CV2021

APES: Audiovisual Person Search in Untrimmed Video

Juan Leon Alcazar, Long Mai, Federico Perazzi +4

Humans are arguably one of the most important subjects in video streams, many real-world applications such as video summarization or video editing workflows often require the autom…

cs.CV2021

MAAS: Multi-modal Assignation for Active Speaker Detection

Juan León-Alcázar, Fabian Caba Heilbron, Ali Thabet +1

Active speaker detection requires a solid integration of multi-modal cues. While individual modalities can approximate a solution, accurate predictions can only be achieved by expl…

cs.CV20206 cited

Real-time Semantic Segmentation with Fast Attention

Ping Hu, Federico Perazzi, Fabian Caba Heilbron +4

In deep CNN based models for semantic segmentation, high accuracy relies on rich spatial context (large receptive fields) and fine spatial details (high resolution), both of which…

cs.CV2020

Active Speakers in Context

Juan Leon Alcazar, Fabian Caba Heilbron, Long Mai +4

Current methods for active speak er detection focus on modeling short-term audiovisual information from a single speaker. Although this strategy can be enough for addressing single…