dialogue systems 1emotional support 1external memory 1large language models 1multi-turn interaction 1situated AI 1
From the 1 of 5 linked papers with an AI index.
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2026
Effective User-defined Keyword Spotting with Dual-stage Matching, Multi-modal Enrollment, and Continual Adaptation
Zhiqi Ai, Han Cheng, Shiyi Mu +3
User-defined keyword spotting (KWS) is crucial for personalized voice interaction, yet existing methods face several challenges: (1) insufficient discriminability among confusable…
eess.AS2024
Enhancing Open-Set Speaker Identification through Rapid Tuning with Speaker Reciprocal Points and Negative Sample
Zhiyong Chen, Zhiqi Ai, Xinnuo Li +1
This paper introduces a novel framework for open-set speaker identification in household environments, playing a crucial role in facilitating seamless human-computer interactions.…
eess.AS2024
StyleFusion TTS: Multimodal Style-control and Enhanced Feature Fusion for Zero-shot Text-to-speech Synthesis
Zhiyong Chen, Xinnuo Li, Zhiqi Ai +1
We introduce StyleFusion-TTS, a prompt and/or audio referenced, style and speaker-controllable, zero-shot text-to-speech (TTS) synthesis system designed to enhance the editability…