3 papers
cs.CL2026
Correcting Gradient-Based Circuit Localization via Interaction-Aware Backpropagation
Joakim Edin, Casper L. Christensen, Róbert Csordás +5
Circuit localization methods aim to identify the subset of model components responsible for specific behaviors in large language models, enabling detailed mechanistic analysis. Mos…
cs.CL2026
A Text-To-Text Alignment Algorithm for Better Evaluation of Modern Speech Recognition Systems
Lasse Borgholt, Jakob Havtorn, Christian Igel +2
Modern neural networks have greatly improved performance across speech recognition benchmarks. However, gains are often driven by frequent words with limited semantic weight, which…
cs.LG2025
Normalized AOPC: Fixing Misleading Faithfulness Metrics for Feature Attribution Explainability
Joakim Edin, Andreas Geert Motzfeldt, Casper L. Christensen +3
Deep neural network predictions are notoriously difficult to interpret. Feature attribution methods aim to explain these predictions by identifying the contribution of each input f…