2 papers
eess.AS2026
No Verifiable Reward for Prosody: Toward Preference-Guided Prosody Learning in TTS
Seungyoun Shin, Dongha Ahn, Jiwoo Kim +1
Recent work reports gains in neural text-to-speech (TTS) with Group Relative Policy Optimization (GRPO). However, in the absence of a verifiable reward for \textit{prosody}, GRPO t…
cs.SD2025
Note-Level Singing Melody Transcription for Time-Aligned Musical Score Generation
Leekyung Kim, Sungwook Jeon, Wan Heo +1
Automatic music transcription converts audio recordings into symbolic representations, facilitating music analysis, retrieval, and generation. A musical note is characterized by pi…