20 citations · 44 across the 20 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2026
Listen, Look, Drive: Coupling Audio Instructions for User-aware VLA-based Autonomous Driving
Ziang Guo, Feng Yang, Xuefeng Zhang +6
Vision Language Action (VLA) models promise an open-vocabulary interface that can translate perceptual ambiguity into semantically grounded driving decisions, yet they still treat…
eess.AS2024
NTC-KWS: Noise-aware CTC for Robust Keyword Spotting
Yu Xi, Haoyu Li, Hao Li +4
In recent years, there has been a growing interest in designing small-footprint yet effective Connectionist Temporal Classification based keyword spotting (CTC-KWS) systems. They a…
eess.AS2024
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech
Yu Xi, Baochen Yang, Hao Li +2
Customizable keyword spotting (KWS) in continuous speech has attracted increasing attention due to its real-world application potential. While contrastive learning (CL) has been wi…