Showing q-bio.NCShow all
2 papers · 1 filter
q-bio.NC2025
Quantifying Uncertainty in Error Consistency: Towards Reliable Behavioral Comparison of Classifiers
Thomas Klein, Sascha Meyen, Wieland Brendel +2
Benchmarking models is a key factor for the rapid progress in machine learning (ML) research. Thus, further progress depends on improving benchmarking metrics. A standard metric to…
q-bio.NC2024
How Aligned are Different Alignment Metrics?
Jannis Ahlert, Thomas Klein, Felix Wichmann +1
In recent years, various methods and benchmarks have been proposed to empirically evaluate the alignment of artificial neural networks to human neural and behavioral data. But how…