2 papers
cs.CL2023
Efficient Spoken Language Recognition via Multilabel Classification
Oriol Nieto, Zeyu Jin, Franck Dernoncourt +1
Spoken language recognition (SLR) is the task of automatically identifying the language present in a speech signal. Existing SLR models are either too computationally expensive or…
eess.AS2022
Audio Similarity is Unreliable as a Proxy for Audio Quality
Pranay Manocha, Zeyu Jin, Adam Finkelstein
Many audio processing tasks require perceptual assessment. However, the time and expense of obtaining ``gold standard'' human judgments limit the availability of such data. Most ap…