From the 2 of 6 linked papers with an AI index.
4 papers · 1 filter
Cover First, Disagree Softly: Rethinking Mismatch-First Active Learning for Frame-Level Audio Classification
Shiqi Zhang, Tuomas Virtanen
The paper studies active learning for frame‑level sound event detection and shows that the common mismatch‑first farthest‑traversal strategy performs poorly under limited labeling…
Greedy Volume Maximization of Gradient Embeddings for Long-Tailed Frame-Level Bioacoustic Active Learning
Shiqi Zhang, Marius FaiÃ, Ariana Strandburg-Peshkin +1
The paper introduces BADGE‑Greedy‑DPP, a deterministic batch selection method that greedily maximizes the volume of gradient embeddings to improve active learning for sparse, long‑…
Mixture-Constrained Max Pooling Improves Separation-Based Bird Species Classification
Yuzhu Wang, Kalle Lahtinen, Patrik Lauha +4
Bird species classification from field recordings remains challenging due to overlapping vocalizations and incomplete species labels. We study source separation as a preprocessing…
Learning Input-Channel Permutation Equivariance for Multi-Channel Source Separation: Reducing Bleeding in Small Music Ensembles
Ruchi Pandey, Jaime Garcia-Martinez, Pablo Cabanas-Molero +5
Microphone bleed is a persistent challenge in small ensembles and orchestral recordings, where close microphones intended for individual instruments also capture leakage from nearb…