audio-visual speech enhancement 1large language models 1reinforcement learning 1reward modeling 1speech quality assessment 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.SD2026
LLM-Guided Reinforcement Learning for Audio-Visual Speech Enhancement
Chih-Ning Chen, Jen-Cheng Hou, Hsin-Min Wang +3
The paper introduces a reinforcement learning framework for audio‑visual speech enhancement that uses a large language model to generate natural‑language feedback, which is convert…
eess.AS2025
Leveraging Self-Supervised Audio-Visual Pretrained Models to Improve Vocoded Speech Intelligibility in Cochlear Implant Simulation
Richard Lee Lai, Jen-Cheng Hou, I-Chun Chern +6
Individuals with hearing impairments face challenges in their ability to comprehend speech, particularly in noisy environments. The aim of this study is to explore the effectivenes…
cs.SD2025
Bridging The Multi-Modality Gaps of Audio, Visual and Linguistic for Speech Enhancement
Meng-Ping Lin, Jen-Cheng Hou, Chia-Wei Chen +4
Speech enhancement (SE) aims to improve the quality and intelligibility of speech in noisy environments. Recent studies have shown that incorporating visual cues in audio signal pr…