2 papers
cs.LG2026
Reading Calibrated Uncertainty from Language Model Trajectories
Aliai Eusebi, Alexander Herzog, Xiaoyu Liang +3
The maximum softmax probability (MSP) represents a default approach when evaluating uncertainty quantification for language model generation with structured output. Although cheap,…
cs.CR2026
On the Reliability and Stability of Selective Methods in Malware Classification Tasks
Alexander Herzog, Aliai Eusebi, Lorenzo Cavallaro
The performance figures of modern drift-adaptive malware classifiers appear promising, but does this translate to genuine operational reliability? The standard evaluation paradigm…