2 papers
cs.RO2026
The Open Ant: A Robot Platform for Reinforcement Learning Research
Elena Sorina Lupu, Patrick Spieler, Khurram Javed +4
Reinforcement learning (RL) research has demonstrated success in both physical and simulated domains; however, the predominant methodology remains rooted in simulations. The predom…
cs.LG2026
Extending Differential Temporal Difference Methods for Episodic Problems
Kris De Asis, Mohamed Elsayed, Jiamin He
Differential temporal difference (TD) methods are value-based reinforcement learning algorithms that have been proposed for infinite-horizon problems. They rely on reward centering…