17 citations · 33 across the 14 of their papers we have counts for
6 papers · 1 filter
Spontaneous Informal Speech Dataset for Punctuation Restoration
Xing Yi Liu, Homayoon Beigi
Presently, punctuation restoration models are evaluated almost solely on well-structured, scripted corpora. On the other hand, real-world ASR systems and post-processing pipelines…
Robust Open-Set Spoken Language Identification and the CU MultiLang Dataset
Mustafa Eyceoz, Justin Lee, Siddharth Pittie +1
Most state-of-the-art spoken language identification models are closed-set; in other words, they can only output a language label from the set of classes they were trained on. Open…
Efficient Ensemble for Multimodal Punctuation Restoration using Time-Delay Neural Network
Xing Yi Liu, Homayoon Beigi
Punctuation restoration plays an essential role in the post-processing procedure of automatic speech recognition, but model efficiency is a key requirement for this task. To that e…
Modernizing Open-Set Speech Language Identification
Mustafa Eyceoz, Justin Lee, Homayoon Beigi
While most modern speech Language Identification methods are closed-set, we want to see if they can be modified and adapted for the open-set problem. When switching to the open-set…
Automatic Spoken Language Identification using a Time-Delay Neural Network
Benjamin Kepecs, Homayoon Beigi
Closed-set spoken language identification is the task of recognizing the language being spoken in a recorded audio clip from a set of known languages. In this study, a language ide…
Cantonese Automatic Speech Recognition Using Transfer Learning from Mandarin
Bryan Li, Xinyue Wang, Homayoon Beigi
We propose a system to develop a basic automatic speech recognizer(ASR) for Cantonese, a low-resource language, through transfer learning of Mandarin, a high-resource language. We…