4 papers
Multi-task self-supervised learning for Robust Speech Recognition
Mirco Ravanelli, Jianyuan Zhong, Santiago Pascual +4
Despite the growing interest in unsupervised learning, extracting meaningful knowledge from unlabelled audio remains an open challenge. To take a step in this direction, we recentl…
Induced Inflection-Set Keyword Search in Speech
Oliver Adams, Matthew Wiesner, Jan Trmal +2
We investigate the problem of searching for a lexeme-set in speech by searching for its inflectional variants. Experimental results indicate how lexeme-set search performance chang…
DiPCo -- Dinner Party Corpus
Maarten Van Segbroeck, Ahmed Zaid, Ksenia Kutsenko +7
We present a speech data corpus that simulates a "dinner party" scenario taking place in an everyday home environment. The corpus was created by recording multiple groups of four A…
Using of heterogeneous corpora for training of an ASR system
Jan Trmal, Gaurav Kumar, Vimal Manohar +3
The paper summarizes the development of the LVCSR system built as a part of the Pashto speech-translation system at the SCALE (Summer Camp for Applied Language Exploration) 2015 wo…