Showing cs.SDShow all
3 papers · 1 filter
cs.SD2024
Video-Foley: Two-Stage Video-To-Sound Generation via Temporal Event Condition For Foley Sound
Junwon Lee, Jaekwon Im, Dabin Kim +1
Foley sound synthesis is crucial for multimedia production, enhancing user experience by synchronizing audio and video both temporally and semantically. Recent studies on automatin…
cs.SD2024
DIFFRENT: A Diffusion Model for Recording Environment Transfer of Speech
Jaekwon Im, Juhan Nam
Properly setting up recording conditions, including microphone type and placement, room acoustics, and ambient noise, is essential to obtaining the desired acoustic characteristics…
cs.SD2022
Neural Vocoder Feature Estimation for Dry Singing Voice Separation
Jaekwon Im, Soonbeom Choi, Sangeon Yong +1
Singing voice separation (SVS) is a task that separates singing voice audio from its mixture with instrumental audio. Previous SVS studies have mainly employed the spectrogram mask…