activity
20222026
most citedImproving Children's Speech Recognition by Fine-tuning Self-supervised Adult Speech Representations

2 citations · 3 across the 9 of their papers we have counts for

collaborators
Showing eess.ASShow all

5 papers · 1 filter

eess.AS2026

Mandarin Humorous Homophone Recognition and Disambiguation in Automatic Speech Recognition

Sicheng Jin, Jinghao Chen, Liuheng Zhou +3

Mandarin homophones remain a key challenge to improving automatic speech recognition (ASR) accuracy due to the amount of potential homophones. Mandarin speakers use this feature ca…

eess.AS2026

Using Phonological-Level Wav2Vec2 for Mandarin Automatic Mispronunciation Detection and Diagnosis

Jinghao Chen, Mostafa Shahin, Beena Ahmed

Automatic mispronunciation detection and diagnosis (MDD) plays a crucial role in L2 Mandarin pronunciation learning. While end-to-end (E2E) based MDD methods have substantially imp…

eess.AS2024

Rethinking Mamba in Speech Processing by Self-Supervised Models

Xiangyu Zhang, Jianbo Ma, Mostafa Shahin +2

The Mamba-based model has demonstrated outstanding performance across tasks in computer vision, natural language processing, and speech processing. However, in the realm of speech…

eess.AS2024

Auto-Landmark: Acoustic Landmark Dataset and Open-Source Toolkit for Landmark Extraction

Xiangyu Zhang, Daijiao Liu, Tianyi Xiao +5

In the speech signal, acoustic landmarks identify times when the acoustic manifestations of the linguistically motivated distinctive features are most salient. Acoustic landmarks h…

eess.AS2023

Phonological Level wav2vec2-based Mispronunciation Detection and Diagnosis Method

Mostafa Shahin, Julien Epps, Beena Ahmed

The automatic identification and analysis of pronunciation errors, known as Mispronunciation Detection and Diagnosis (MDD) plays a crucial role in Computer Aided Pronunciation Lear…