activity
20172023
most citedHoudini: Fooling Deep Structured Prediction Models

164 citations · 167 across the 9 of their papers we have counts for

collaborators
Showing eess.ASShow all

6 papers · 1 filter

eess.AS2023

Open-vocabulary Keyword-spotting with Adaptive Instance Normalization

Aviv Navon, Aviv Shamsian, Neta Glazer +2

Open vocabulary keyword spotting is a crucial and challenging task in automatic speech recognition (ASR) that focuses on detecting user-defined keywords within a spoken utterance.…

eess.AS2021

CNN-based Spoken Term Detection and Localization without Dynamic Programming

Tzeviya Sylvia Fuchs, Yael Segal, Joseph Keshet

In this paper, we propose a spoken term detection algorithm for simultaneous prediction and localization of in-vocabulary and out-of-vocabulary terms within an audio segment. The p…

eess.AS2020

Self-Supervised Contrastive Learning for Unsupervised Phoneme Segmentation

Felix Kreuk, Joseph Keshet, Yossi Adi

We propose a self-supervised representation learning model for the task of unsupervised phoneme boundary detection. The model is a convolutional neural network that operates direct…

eess.AS2020

Phoneme Boundary Detection using Learnable Segmental Features

Felix Kreuk, Yaniv Sheena, Joseph Keshet +1

Phoneme boundary detection plays an essential first step for a variety of speech processing applications such as speaker diarization, speech science, keyword spotting, etc. In this…

eess.AS20192 cited

Dr.VOT : Measuring Positive and Negative Voice Onset Time in the Wild

Yosi Shrem, Matthew Goldrick, Joseph Keshet

Voice Onset Time (VOT), a key measurement of speech for basic research and applied medical studies, is the time between the onset of a stop burst and the onset of voicing. When the…

eess.AS2019

SpeechYOLO: Detection and Localization of Speech Objects

Yael Segal, Tzeviya Sylvia Fuchs, Joseph Keshet

In this paper, we propose to apply object detection methods from the vision domain on the speech recognition domain, by treating audio fragments as objects. More specifically, we p…