14 papers
An evaluation of data augmentation methods for sound scene geotagging
Helen L. Bear, Veronica Morfi, Emmanouil Benetos
Sound scene geotagging is a new topic of research which has evolved from acoustic scene classification. It is motivated by the idea of audio surveillance. Not content with only des…
Visually Exploring Multi-Purpose Audio Data
David Heise, Helen L. Bear
We analyse multi-purpose audio using tools to visualise similarities within the data that may be observed via unsupervised methods. The success of machine learning classifiers is a…
Memory Controlled Sequential Self Attention for Sound Recognition
Arjun Pankajakshan, Helen L. Bear, Vinod Subramanian +1
In this paper we investigate the importance of the extent of memory in sequential self attention for sound recognition. We propose to use a memory controlled sequential self attent…
Alternative Visual Units for an Optimized Phoneme-Based Lipreading System
Helen Bear, Richard Harvey
Lipreading is understanding speech from observed lip movements. An observed series of lip motions is an ordered sequence of visual lip gestures. These gestures are commonly known,…
Polyphonic Sound Event and Sound Activity Detection: A Multi-task approach
Arjun Pankajakshan, Helen L. Bear, Emmanouil Benetos
Polyphonic Sound Event Detection (SED) in real-world recordings is a challenging task because of the dynamic polyphony level, intensity, and duration of sound events. Current polyp…
City classification from multiple real-world sound scenes
Helen L. Bear, Toni Heittola, Annamaria Mesaros +2
The majority of sound scene analysis work focuses on one of two clearly defined tasks: acoustic scene classification or sound event detection. Whilst this separation of tasks is us…