5 papers
CounterFlow: A Two-Phase Inference-Time Sampling for Counterfactual Video Foley Generation
Gyubin Lee, Junwon Lee, Juhan Nam
We investigate Counterfactual Video Foley Generation, which aims to adopt a sound-source identity that contradicts the visual evidence while remaining temporally synchronized to a…
PiAnnotate: A Web Annotation Tool for Piano Fingering, with a Diagnostic Probe
Joonhyung Bae, Kirak Kim, Hyeyoon Cho +8
Piano fingering shapes how a passage can be played, yet it is difficult to label after a performance. An annotator must decide which finger produced each note while reconciling the…
Tipiano: Cascaded Piano Hand Motion Synthesis via Fingertip Priors
Joonhyung Bae, Kirak Kim, Hyeyoon Cho +9
Synthesizing realistic piano hand motions requires both precision and naturalness. Physics-based methods achieve precision but produce stiff motions; data-driven models learn natur…
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
Gyubin Lee, Truong Nhat Nguyen Bao, Jaesik Yoon +4
Diffusion models have demonstrated strong generative capabilities across domains ranging from image synthesis to complex reasoning tasks. However, most inference-time scaling metho…
CONMOD: Controllable Neural Frame-based Modulation Effects
Gyubin Lee, Hounsu Kim, Junwon Lee +1
Deep learning models have seen widespread use in modelling LFO-driven audio effects, such as phaser and flanger. Although existing neural architectures exhibit high-quality emulati…