Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Beyond Word Error Rate: Auditing the Diversity Tax in Speech Recognition through Dataset Cartography
Ting-Hui Cheng, Line H. Clemmensen, Sneha Das
Automatic speech recognition (ASR) systems are predominantly evaluated using the Word Error Rate (WER). However, raw token-level metrics fail to capture semantic fidelity and routi…
cs.LG2026
Intra-Fairness Dynamics: The Bias Spillover Effect in Targeted LLM Alignment
Eva Paraschou, Line Harder Clemmensen, Sneha Das
Conventional large language model (LLM) fairness alignment largely focuses on mitigating bias along single sensitive attributes, overlooking fairness as an inherently multidimensio…
cs.LG2025
Evaluation of Stress Detection as Time Series Events -- A Novel Window-Based F1-Metric
Harald Vilhelm Skat-Rørdam, Sneha Das, Kathrine Sofie Rasmussen +2
Accurate evaluation of event detection in time series is essential for applications such as stress monitoring with wearable devices, where ground truth is typically annotated as si…