Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Human-Alignment, Calibration, and Activation Patterns in Large Language Model Uncertainty
Kyle Moore, Jesse Roberts, Daryl Watson +2
Uncertainty Quantification is a large and growing subfield of large language model behavioral analysis. Primarily to recognize and combat hallucination, the field has largely focus…
cs.CL2026
KARMA: Karma-Aligned Reward Model Adaptation
Jared Scott, Jesse Roberts
Human communication depends on implicit social signals where effectiveness is shaped by tone, context, and conversational norms rather than semantic content alone. We introduce KAR…