papers

Publications (11)

cs.CL2026

Arab Voices: Mapping Standard and Dialectal Arabic Speech Technology

Peter Sullivan, AbdelRahim Elmadany, Alcides Alcoba Inciarte +1

Dialectal Arabic (DA) speech data vary widely in domain coverage, dialect labeling practices, and recording conditions, complicating cross-dataset comparison and model evaluation.…

cs.CL2025

Quantum NLP models on Natural Language Inference

Ling Sun, Peter Sullivan, Michael Martin +1

Quantum natural language processing (QNLP) offers a novel approach to semantic modeling by embedding compositional structure directly into quantum circuits. This paper investigates…

cs.CL2025

NADI 2025: The First Multidialectal Arabic Speech Processing Shared Task

Bashar Talafha, Hawau Olamide Toyin, Peter Sullivan +9

We present the findings of the sixth Nuanced Arabic Dialect Identification (NADI 2025) Shared Task, which focused on Arabic speech dialect processing across three subtasks: spoken…

cs.CL2023

VoxArabica: A Robust Dialect-Aware Arabic Speech Recognition System

Abdul Waheed, Bashar Talafha, Peter Sullivan +2

Arabic is a complex language with many varieties and dialects spoken by over 450 millions all around the world. Due to the linguistic diversity and variations, it is challenging to…

cs.SD2025

On Barriers to Archival Audio Processing

Peter Sullivan, Muhammad Abdul-Mageed

In this study, we leverage a unique UNESCO collection of mid-20th century radio recordings to probe the robustness of modern off-the-shelf language identification (LID) and speaker…

eess.AS2021

Speech Technology for Everyone: Automatic Speech Recognition for Non-Native English with Transfer Learning

Toshiko Shibano, Xinyi Zhang, Mia Taige Li +3

To address the performance gap of English ASR models on L2 English speakers, we evaluate fine-tuning of pretrained wav2vec 2.0 models (Baevski et al., 2020; Xu et al., 2021) on L2-…