activity
20222024
most citedPolySpeech: Exploring Unified Multitask Speech Models for Competitiveness with Single-task Models

1 citations · 1 across the 5 of their papers we have counts for

collaborators

5 papers

eess.AS2024

On Calibration of Speech Classification Models: Insights from Energy-Based Model Investigations

Yaqian Hao, Chenguang Hu, Yingying Gao +2

For speech classification tasks, deep learning models often achieve high accuracy but exhibit shortcomings in calibration, manifesting as classifiers exhibiting overconfidence. The…

eess.AS2024

CEC: A Noisy Label Detection Method for Speaker Recognition

Yao Shen, Yingying Gao, Yaqian Hao +4

Noisy labels are inevitable, even in well-annotated datasets. The detection of noisy labels is of significant importance to enhance the robustness of speaker recognition models. In…

cs.CL20241 cited

PolySpeech: Exploring Unified Multitask Speech Models for Competitiveness with Single-task Models

Runyan Yang, Huibao Yang, Xiqing Zhang +6

Recently, there have been attempts to integrate various speech processing tasks into a unified model. However, few previous works directly demonstrated that joint optimization of d…

cs.SD2023

VE-KWS: Visual Modality Enhanced End-to-End Keyword Spotting

Ao Zhang, He Wang, Pengcheng Guo +5

The performance of the keyword spotting (KWS) system based on audio modality, commonly measured in false alarms and false rejects, degrades significantly under the far field and no…

eess.AS2022

Meta Auxiliary Learning for Low-resource Spoken Language Understanding

Yingying Gao, Junlan Feng, Chao Deng +1

Spoken language understanding (SLU) treats automatic speech recognition (ASR) and natural language understanding (NLU) as a unified task and usually suffers from data scarcity. We…