activity
20232026
most citedSpeechPrompt: Prompting Speech Language Models for Speech Processing Tasks

8 citations · 20 across the 23 of their papers we have counts for

collaborators
Showing 2026 · eess.ASShow all

5 papers · 2 filters

eess.AS2026

MP-Bench: Evaluating Voice Agents as a Multiparty Conversation Participant

Yi-Jen Shih, Shih-Yun Shan Kuan, Guan-Ting Lin +7

Conversational voice agents have advanced significantly, offering increasingly natural human-machine interactions through both cascaded and end-to-end architectures. However, while…

eess.AS2026

AudioICL-Bench: A Benchmark for Large Audio Language Model In-Context Learning

Jia-Hung Chen, Yi-Cheng Lin, Kai-Wei Chang +2

In-context learning (ICL) promises training-free adaptation for audio, where labeling every new condition is costly. Yet existing audio ICL studies largely measure Task Recognition…

eess.AS2026

Unsupervised Speech Recognition at the Syllable Level

Liming Wang, Kai-Wei Chang, Kunio Kashino +3

Training speech recognizers with unpaired speech and text -- known as unsupervised speech recognition (UASR) -- is a crucial step toward extending ASR to low-resource languages in…

eess.AS2026

RRP-Voice: A Longitudinal Dataset and Benchmark for Recurrent Respiratory Papillomatosis Detection

Wenze Ren, Ke-Han Lu, Kai-Wei Chang +8

Deep learning has advanced pathological voice detection rapidly, yet rare laryngeal diseases remain underexplored due to data scarcity. Recurrent Respiratory Papillomatosis (RRP) e…

eess.AS2026

AQAScore: Evaluating Semantic Alignment in Text-to-Audio Generation via Audio Question Answering

Chun-Yi Kuan, Kai-Wei Chang, Hung-yi Lee

Although text-to-audio generation has made remarkable progress in realism and diversity, the development of evaluation metrics has not kept pace. Widely-adopted approaches, typical…