1 paper · 1 filter
Adrian S. Roman, Aiden Chang, Gerardo Meza +1
We present SELDVisualSynth, a tool for generating synthetic videos for audio-visual sound event localization and detection (SELD). Our approach incorporates real-world background i…