1 paper · 1 filter
Han Wang, Philippe Beardsell, Boning Li +4
Reasoning in large language models (LLMs) is often grounded in human text, human demonstrations, and human-generated rationales. For equilibrium reasoning in complex games, however…