Showing eess.ASShow all
2 papers · 1 filter
eess.AS2026
ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment
Jun-Hak Yun, Seung-Bin Kim, Seong-Whan Lee
Recent advancements in text-guided audio generation have yielded promising results in diverse domains, including sound effects, speech, and music. However, jointly generating speec…
eess.AS2025
FLowHigh: Towards Efficient and High-Quality Audio Super-Resolution with Single-Step Flow Matching
Jun-Hak Yun, Seung-Bin Kim, Seong-Whan Lee
Audio super-resolution is challenging owing to its ill-posed nature. Recently, the application of diffusion models in audio super-resolution has shown promising results in alleviat…