2 papers
cs.LG2025
4Hammer: a board-game reinforcement learning environment for the hour long time frame
Massimo Fioravanti, Giovanni Agosta
Large Language Models (LLMs) have demonstrated strong performance on tasks with short time frames, but struggle with tasks requiring longer durations. While datasets covering exten…
cs.PL2025
Rulebook: bringing co-routines to reinforcement learning environments
Massimo Fioravanti, Samuele Pasini, Giovanni Agosta
Reinforcement learning (RL) algorithms, due to their reliance on external systems to learn from, require digital environments (e.g., simulators) with very simple interfaces, which…