37 citations · 37 across the 1 of their papers we have counts for
2 papers
cs.CV2018
Extracting textual overlays from social media videos using neural networks
Adam Słucki, Tomasz Trzcinski, Adam Bielski +1
Textual overlays are often used in social media videos as people who watch them without the sound would otherwise miss essential information conveyed in the audio stream. This is w…
cs.SD2017★ 37 cited
Speaker Diarization using Deep Recurrent Convolutional Neural Networks for Speaker Embeddings
Pawel Cyrta, Tomasz Trzciński, Wojciech Stokowiec
In this paper we propose a new method of speaker diarization that employs a deep learning architecture to learn speaker embeddings. In contrast to the traditional approaches that b…