4 papers
Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility
Michael A. Lepori, Jennifer Hu, Ishita Dasgupta +3
Language models (LMs) are used for a diverse range of tasks, from question answering to writing fantastical stories. In order to reliably accomplish these tasks, LMs must be able t…
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models
Roberto-Rafael Maura-Rivero, Chirag Nagpal, Roma Patel +1
Current methods that train large language models (LLMs) with reinforcement learning feedback, often resort to averaging outputs of multiple rewards functions during training. This…
Steering Language Models with Game-Theoretic Solvers
Ian Gemp, Roma Patel, Yoram Bachrach +6
Mathematical models of interactions among rational agents have long been studied in game theory. However these interactions are often over a small set of discrete game actions whic…
Skill Generalization with Verbs
Rachel Ma, Lyndon Lam, Benjamin A. Spiegel +6
It is imperative that robots can understand natural language commands issued by humans. Such commands typically contain verbs that signify what action should be performed on a give…