Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Entropy-informed Decoding: Adaptive Information-Driven Branching
Benjamin Patrick Evans, Sumitra Ganesh, Leo Ardon
Large language models (LLMs) achieve remarkable generative performance, yet their output quality is dependent on the decoding strategy. While sampling-based methods (e.g., top-k, n…
cs.LG2025
Learning in Stackelberg Mean Field Games: A Non-Asymptotic Analysis
Sihan Zeng, Benjamin Patrick Evans, Sujay Bhatt +3
We study policy optimization in Stackelberg mean field games (MFGs), a hierarchical framework for modeling the strategic interaction between a single leader and an infinitely large…
cs.LG2025
Modelling bounded rational decision-making through Wasserstein constraints
Benjamin Patrick Evans, Leo Ardon, Sumitra Ganesh
Modelling bounded rational decision-making through information constrained processing provides a principled approach for representing departures from rationality within a reinforce…