3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.AI2024★ 3 cited
Planning in Strawberry Fields: Evaluating and Improving the Planning and Scheduling Capabilities of LRM o1
Karthik Valmeekam, Kaya Stechly, Atharva Gundawar +1
The ability to plan a course of action that achieves a desired state of affairs has long been considered a core competence of intelligent agents and has been an integral part of AI…
cs.AI2024
Superior Computer Chess with Model Predictive Control, Reinforcement Learning, and Rollout
Atharva Gundawar, Yuchao Li, Dimitri Bertsekas
In this paper we apply model predictive control (MPC), rollout, and reinforcement learning (RL) methodologies to computer chess. We introduce a new architecture for move selection,…