Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
From token probabilities to calibrated confidence: An empirical study of mathematical question answering
Avery Ma, Lorne Schell, Vin Bhaskara +1
Confidence estimation for large language models (LLMs) aims to estimate the probability that a generated answer is correct, while calibration aligns these estimates with empirical…
cs.LG2024
Improving Adversarial Transferability via Model Alignment
Avery Ma, Amir-massoud Farahmand, Yangchen Pan +2
Neural networks are susceptible to adversarial perturbations that are transferable across different models. In this paper, we introduce a novel model alignment technique aimed at i…