4 papers
RLZero: Direct Policy Inference from Language Without In-Domain Supervision
Harshit Sikchi, Siddhant Agarwal, Pranaya Jajoo +6
The reward hypothesis states that all goals and purposes can be understood as the maximization of a received scalar reward signal. However, in practice, defining such a reward sign…
DreamGarden: A Designer Assistant for Growing Games from a Single Prompt
Sam Earle, Samyak Parajuli, Andrzej Banburski-Fahey
Coding assistants are increasingly leveraged in game design, both generating code and making high-level plans. To what degree can these tools align with developer workflows, and wh…
AMAGO-2: Breaking the Multi-Task Barrier in Meta-Reinforcement Learning with Transformers
Jake Grigsby, Justin Sasek, Samyak Parajuli +3
Language models trained on diverse datasets unlock generalization by in-context learning. Reinforcement Learning (RL) policies can achieve a similar effect by meta-learning within…
Social Conjuring: Multi-User Runtime Collaboration with AI in Building Virtual 3D Worlds
Amina Kobenova, Cyan DeVeaux, Samyak Parajuli +3
Generative artificial intelligence has shown promise in prompting virtual worlds into existence, yet little attention has been given to understanding how this process unfolds as so…