Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
GSRM: Generative Speech Reward Model for Speech RLHF
Maohao Shen, Tejas Jayashankar, Osama Hanna +10
Recent advances in speech language models, such as GPT-4o Voice Mode and Gemini Live, have demonstrated promising speech generation capabilities. Nevertheless, the aesthetic natura…
cs.SD2025
Directional Source Separation for Robust Speech Recognition on Smart Glasses
Tiantian Feng, Ju Lin, Yiteng Huang +7
Modern smart glasses leverage advanced audio sensing and machine learning technologies to offer real-time transcribing and captioning services, considerably enriching human experie…