Showing eess.ASShow all
2 papers · 1 filter
eess.AS2026
Cross-Modal Bottleneck Fusion For Noise Robust Audio-Visual Speech Recognition
Seaone Ok, Min Jun Choi, Eungbeom Kim +2
Audio-Visual Speech Recognition (AVSR) leverages both acoustic and visual cues to improve speech recognition under noisy conditions. A central question is how to design a fusion me…
eess.AS2024
Differentiable Modal Synthesis for Physical Modeling of Planar String Sound and Motion Simulation
Jin Woo Lee, Jaehyun Park, Min Jun Choi +1
While significant advancements have been made in music generation and differentiable sound synthesis within machine learning and computer audition, the simulation of instrument vib…