3 papers
cs.SD2026
BAT: Better Audio Transformer Guided by Convex Gated Probing
Houtan Ghaffari, Lukas Rauch, Christoph Scholz +1
Probing is widely adopted in computer vision to faithfully evaluate self-supervised learning (SSL) embeddings, as finetuning may misrepresent their inherent quality. In contrast, a…
cs.LG2025
Data-Efficient Self-Supervised Algorithms for Fine-Grained Birdsong Analysis
Houtan Ghaffari, Lukas Rauch, Paul Devos
Research in bioacoustics, neuroscience, and linguistics often uses birdsong as a proxy to acquire knowledge across diverse areas. This requires audio models to annotate and parse t…
cs.LG2024
Comparison of self-supervised in-domain and supervised out-domain transfer learning for bird species recognition
Houtan Ghaffari, Paul Devos
Transferring the weights of a pre-trained model to assist another task has become a crucial part of modern deep learning, particularly in data-scarce scenarios. Pre-training refers…