3 papers
eess.AS2026
Cross-Modal Bottleneck Fusion For Noise Robust Audio-Visual Speech Recognition
Seaone Ok, Min Jun Choi, Eungbeom Kim +2
Audio-Visual Speech Recognition (AVSR) leverages both acoustic and visual cues to improve speech recognition under noisy conditions. A central question is how to design a fusion me…
cs.SD2025
Differentiable Acoustic Radiance Transfer
Sungho Lee, Matteo Scerbo, Seungu Han +3
Geometric acoustics is an efficient framework for room acoustics modeling, governed by the canonical time-dependent rendering equation. Acoustic radiance transfer (ART) solves the…
eess.AS2024
Differentiable Modal Synthesis for Physical Modeling of Planar String Sound and Motion Simulation
Jin Woo Lee, Jaehyun Park, Min Jun Choi +1
While significant advancements have been made in music generation and differentiable sound synthesis within machine learning and computer audition, the simulation of instrument vib…