164 citations · 167 across the 9 of their papers we have counts for
6 papers · 1 filter
Open-vocabulary Keyword-spotting with Adaptive Instance Normalization
Aviv Navon, Aviv Shamsian, Neta Glazer +2
Open vocabulary keyword spotting is a crucial and challenging task in automatic speech recognition (ASR) that focuses on detecting user-defined keywords within a spoken utterance.…
CNN-based Spoken Term Detection and Localization without Dynamic Programming
Tzeviya Sylvia Fuchs, Yael Segal, Joseph Keshet
In this paper, we propose a spoken term detection algorithm for simultaneous prediction and localization of in-vocabulary and out-of-vocabulary terms within an audio segment. The p…
Self-Supervised Contrastive Learning for Unsupervised Phoneme Segmentation
Felix Kreuk, Joseph Keshet, Yossi Adi
We propose a self-supervised representation learning model for the task of unsupervised phoneme boundary detection. The model is a convolutional neural network that operates direct…
Phoneme Boundary Detection using Learnable Segmental Features
Felix Kreuk, Yaniv Sheena, Joseph Keshet +1
Phoneme boundary detection plays an essential first step for a variety of speech processing applications such as speaker diarization, speech science, keyword spotting, etc. In this…
Dr.VOT : Measuring Positive and Negative Voice Onset Time in the Wild
Yosi Shrem, Matthew Goldrick, Joseph Keshet
Voice Onset Time (VOT), a key measurement of speech for basic research and applied medical studies, is the time between the onset of a stop burst and the onset of voicing. When the…
SpeechYOLO: Detection and Localization of Speech Objects
Yael Segal, Tzeviya Sylvia Fuchs, Joseph Keshet
In this paper, we propose to apply object detection methods from the vision domain on the speech recognition domain, by treating audio fragments as objects. More specifically, we p…