2 citations · 2 across the 5 of their papers we have counts for
5 papers
Qomhra: A Bilingual Irish and English Large Language Model
Joseph McInerney, Khanh-Tung Tran, Liam Lonergan +3
Large language model (LLM) research and development has overwhelmingly focused on the world's major languages, leading to under-representation of low-resource languages such as Iri…
Fotheidil: an Automatic Transcription System for the Irish Language
Liam Lonergan, Ibon Saratxaga, John Sloan +5
This paper sets out the first web-based transcription system for the Irish language - Fotheidil, a system that utilises speech-related AI technologies as part of the ABAIR initiati…
Low-resource speech recognition and dialect identification of Irish in a multi-task framework
Liam Lonergan, Mengjie Qian, Neasa Ní Chiaráin +2
This paper explores the use of Hybrid CTC/Attention encoder-decoder models trained with Intermediate CTC (InterCTC) for Irish (Gaelic) low-resource speech recognition (ASR) and dia…
Towards spoken dialect identification of Irish
Liam Lonergan, Mengjie Qian, Neasa Ní Chiaráin +2
The Irish language is rich in its diversity of dialects and accents. This compounds the difficulty of creating a speech recognition system for the low-resource language, as such a…
Towards dialect-inclusive recognition in a low-resource language: are balanced corpora the answer?
Liam Lonergan, Mengjie Qian, Neasa Ní Chiaráin +2
ASR systems are generally built for the spoken 'standard', and their performance declines for non-standard dialects/varieties. This is a problem for a language like Irish, where th…