3 citations · 3 across the 8 of their papers we have counts for
8 papers
Generating Diverse Audio-Visual 360 Soundscapes for Sound Event Localization and Detection
Adrian S. Roman, Aiden Chang, Gerardo Meza +1
We present SELDVisualSynth, a tool for generating synthetic videos for audio-visual sound event localization and detection (SELD). Our approach incorporates real-world background i…
Design and Implementation of the Transparent, Interpretable, and Multimodal (TIM) AR Personal Assistant
Erin McGowan, Joao Rulff, Sonia Castelo +11
The concept of an AI assistant for task guidance is rapidly shifting from a science fiction staple to an impending reality. Such a system is inherently complex, requiring models fo…
Analyzing Pitch Content in Traditional Ghanaian Seperewa Songs
Kelvin L Walls, Iran R Roman, Kelsey Van Ert +2
This study examines the pitch content in traditional Ghanaian seperewa (Akan harp-lute) songs, utilizing a unique dataset from field recordings of the mid-twentieth century. We sel…
HuBar: A Visual Analytics Tool to Explore Human Behaviour based on fNIRS in AR guidance systems
Sonia Castelo, Joao Rulff, Parikshit Solunke +11
The concept of an intelligent augmented reality (AR) assistant has significant, wide-ranging applications, with potential uses in medicine, military, and mechanics domains. Such an…
Spatial Scaper: A Library to Simulate and Augment Soundscapes for Sound Event Localization and Detection in Realistic Rooms
Iran R. Roman, Christopher Ick, Sivan Ding +3
Sound event localization and detection (SELD) is an important task in machine listening. Major advancements rely on simulated data with sound events in specific rooms and strong sp…
Robust DOA estimation using deep acoustic imaging
Adrian S. Roman, Iran R. Roman, Juan P. Bello
Direction of arrival estimation (DoAE) aims at tracking a sound in azimuth and elevation. Recent advancements include data-driven models with inputs derived from ambisonics intensi…