5 papers
VoiceMorph: How AI Voice Morphing Reveals the Boundaries of Auditory Self-Recognition
Kye Shimizu, Minghan Gao, Ananya Ganesh +1
This study investigated auditory self-recognition boundaries using AI voice morphing technology, examining when individuals cease recognizing their own voice. Through controlled mo…
The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition
Ming Gao, Shilong Wu, Hang Chen +6
Meetings are a valuable yet challenging scenario for speech applications due to complex acoustic conditions. This paper summarizes the outcomes of the MISP 2025 Challenge, hosted a…
CDSD: Chinese Dysarthria Speech Database
Yan Wang, Mengyi Sun, Xinchen Kang +4
Dysarthric speech poses significant challenges for individuals with dysarthria, impacting their ability to communicate socially. Despite the widespread use of Automatic Speech Reco…
Exploring the Impact of Emotional Voice Integration in Sign-to-Speech Translators for Deaf-to-Hearing Communication
Hyunchul Lim, Minghan Gao, Franklin Mingzhe Li +4
Emotional voice communication plays a crucial role in effective daily interactions. Deaf and hard-of-hearing (DHH) individuals often rely on facial expressions to supplement sign l…
DocEDA: Automated Extraction and Design of Analog Circuits from Documents with Large Language Model
Hong Cai Chen, Longchang Wu, Ming Gao +3
Efficient and accurate extraction of electrical parameters from circuit datasheets and design documents is critical for accelerating circuit design in Electronic Design Automation…