2 papers
cs.LG2026
From token probabilities to calibrated confidence: An empirical study of mathematical question answering
Avery Ma, Lorne Schell, Vin Bhaskara +1
Confidence estimation for large language models (LLMs) aims to estimate the probability that a generated answer is correct, while calibration aligns these estimates with empirical…
cs.LG2025
AutoGrid AI: Deep Reinforcement Learning Framework for Autonomous Microgrid Management
Kenny Guo, Nicholas Eckhert, Krish Chhajer +2
We present a deep reinforcement learning-based framework for autonomous microgrid management. tailored for remote communities. Using deep reinforcement learning and time-series for…