1 paper · 1 filter
Akshat Singh Jaswal, Ashish Baghel, Paras Chopra
Reinforcement learning systems rely on environment interfaces that specify observations and reward functions, yet constructing these interfaces for new tasks often requires substan…