Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
MAEB: Massive Audio Embedding Benchmark
Adnan El Assadi, Isaac Chung, Chenghao Xiao +15
We introduce the Massive Audio Embedding Benchmark (MAEB), a large-scale benchmark covering 30 tasks across speech, music, environmental sounds, and cross-modal audio-text reasonin…
cs.SD2025
Multi-Stage Speaker Diarization for Noisy Classrooms
Ali Sartaz Khan, Tolulope Ogunremi, Ahmed Adel Attia +1
Speaker diarization, the process of identifying "who spoke when" in audio recordings, is essential for understanding classroom dynamics. However, classroom settings present distinc…