1 citations · 1 across the 12 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
Where Does the Sound Go? Tracing Acoustic Information Loss in Audio-Conditioned LLMs
Song-ha Jo, Sehyun Lee, Soyoon Kim +2
Audio-conditioned language models often underuse acoustic cues such as prosody, emotion, and non-speech sounds, raising the question of whether ASR-supervised frontends discard thi…
cs.SD2025
Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models
Kyowoon Lee, Artyom Stitsyuk, Gunu Jho +2
Recent advances in Text-to-Speech (TTS) have significantly improved speech naturalness, increasing the demand for precise prosody control and mispronunciation correction. Existing…