45 citations · 174 across the 34 of their papers we have counts for
Showing 2023 · eess.ASShow all
2 papers · 2 filters
eess.AS2023
VoiceLDM: Text-to-Speech with Environmental Context
Yeonghyeon Lee, Inmo Yeon, Juhan Nam +1
This paper presents VoiceLDM, a model designed to produce audio that accurately follows two distinct natural language text prompts: the description prompt and the content prompt. T…
eess.AS2023★ 1 cited
All-In-One Metrical And Functional Structure Analysis With Neighborhood Attentions on Demixed Audio
Taejun Kim, Juhan Nam
Music is characterized by complex hierarchical structures. Developing a comprehensive model to capture these structures has been a significant challenge in the field of Music Infor…