Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Provably Optimal Learning Algorithms for Assistance Games
Nivasini Ananthakrishnan, Mark Bedaywi, Michael I. Jordan +2
This paper studies an online variant of the assistance games framework, where an informed agent and an uninformed agent repeatedly interact over timesteps to optimize a common…
cs.LG2024
PID Accelerated Temporal Difference Algorithms
Mark Bedaywi, Amin Rakhsha, Amir-massoud Farahmand
Long-horizon tasks, which have a large discount factor, pose a challenge for most conventional reinforcement learning (RL) algorithms. Algorithms such as Value Iteration and Tempor…