2 citations · 3 across the 9 of their papers we have counts for
5 papers · 1 filter
Mandarin Humorous Homophone Recognition and Disambiguation in Automatic Speech Recognition
Sicheng Jin, Jinghao Chen, Liuheng Zhou +3
Mandarin homophones remain a key challenge to improving automatic speech recognition (ASR) accuracy due to the amount of potential homophones. Mandarin speakers use this feature ca…
Using Phonological-Level Wav2Vec2 for Mandarin Automatic Mispronunciation Detection and Diagnosis
Jinghao Chen, Mostafa Shahin, Beena Ahmed
Automatic mispronunciation detection and diagnosis (MDD) plays a crucial role in L2 Mandarin pronunciation learning. While end-to-end (E2E) based MDD methods have substantially imp…
Rethinking Mamba in Speech Processing by Self-Supervised Models
Xiangyu Zhang, Jianbo Ma, Mostafa Shahin +2
The Mamba-based model has demonstrated outstanding performance across tasks in computer vision, natural language processing, and speech processing. However, in the realm of speech…
Auto-Landmark: Acoustic Landmark Dataset and Open-Source Toolkit for Landmark Extraction
Xiangyu Zhang, Daijiao Liu, Tianyi Xiao +5
In the speech signal, acoustic landmarks identify times when the acoustic manifestations of the linguistically motivated distinctive features are most salient. Acoustic landmarks h…
Phonological Level wav2vec2-based Mispronunciation Detection and Diagnosis Method
Mostafa Shahin, Julien Epps, Beena Ahmed
The automatic identification and analysis of pronunciation errors, known as Mispronunciation Detection and Diagnosis (MDD) plays a crucial role in Computer Aided Pronunciation Lear…