collaborators

6 papers

cs.SD2025

SLM-TTA: A Framework for Test-Time Adaptation of Generative Spoken Language Models

Yuan-Kuei Wu, Yang Liu, Yiteng Huang +9

Spoken Language Models (SLMs) are increasingly central to modern speech-driven applications, but performance degrades under acoustic shift - real-world noise, reverberation, and mi…

eess.AS2025

Multi-Channel Differential ASR for Robust Wearer Speech Recognition on Smart Glasses

Yufeng Yang, Yiteng Huang, Yong Xu +9

With the growing adoption of wearable devices such as smart glasses for AI assistants, wearer speech recognition (WSR) is becoming increasingly critical to next-generation human-co…

eess.AS2025

MMW: Side Talk Rejection Multi-Microphone Whisper on Smart Glasses

Yang Liu, Li Wan, Yiteng Huang +5

Smart glasses are increasingly positioned as the next-generation interface for ubiquitous access to large language models (LLMs). Nevertheless, achieving reliable interaction in re…

cs.SD2025

Directional Source Separation for Robust Speech Recognition on Smart Glasses

Tiantian Feng, Ju Lin, Yiteng Huang +7

Modern smart glasses leverage advanced audio sensing and machine learning technologies to offer real-time transcribing and captioning services, considerably enriching human experie…

eess.AS2024

MASV: Speaker Verification with Global and Local Context Mamba

Yang Liu, Li Wan, Yiteng Huang +3

Deep learning models like Convolutional Neural Networks and transformers have shown impressive capabilities in speech verification, gaining considerable attention in the research c…

cs.CL2024

Query-by-Example Keyword Spotting Using Spectral-Temporal Graph Attentive Pooling and Multi-Task Learning

Zhenyu Wang, Shuyu Kong, Li Wan +6

Existing keyword spotting (KWS) systems primarily rely on predefined keyword phrases. However, the ability to recognize customized keywords is crucial for tailoring interactions wi…