Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
BAT: Better Audio Transformer Guided by Convex Gated Probing
Houtan Ghaffari, Lukas Rauch, Christoph Scholz +1
Probing is widely adopted in computer vision to faithfully evaluate self-supervised learning (SSL) embeddings, as finetuning may misrepresent their inherent quality. In contrast, a…
cs.SD2025
Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio Classification
Lukas Rauch, René Heinrich, Houtan Ghaffari +4
Although probing frozen models has become a standard evaluation paradigm, self-supervised learning in audio defaults to fine-tuning when pursuing state-of-the-art on AudioSet. A ke…