2 papers
cs.CL2025
Emergent World Beliefs: Exploring Transformers in Stochastic Games
Adam Kamel, Tanish Rastogi, Michael Ma +2
Transformer-based large language models (LLMs) have demonstrated strong reasoning abilities across diverse fields, from solving programming challenges to competing in strategy-inte…
cs.AI2025
Automated Reward Design for Gran Turismo
Michel Ma, Takuma Seno, Kaushik Subramanian +3
When designing reinforcement learning (RL) agents, a designer communicates the desired agent behavior through the definition of reward functions - numerical feedback given to the a…