10 citations · 18 across the 5 of their papers we have counts for
3 papers · 1 filter
Minimum-Delay Adaptation in Non-Stationary Reinforcement Learning via Online High-Confidence Change-Point Detection
Lucas N. Alegre, Ana L. C. Bazzan, Bruno C. da Silva
Non-stationary environments are challenging for reinforcement learning algorithms. If the state transition and/or reward functions change based on latent factors, the agent is effe…
Universal Off-Policy Evaluation
Yash Chandak, Scott Niekum, Bruno Castro da Silva +3
When faced with sequential decision-making problems, it is often useful to be able to predict what would happen if decisions were made using a new policy. Those predictions must of…
Optimal Options for Multi-Task Reinforcement Learning Under Time Constraints
Manuel Del Verme, Bruno Castro da Silva, Gianluca Baldassarre
Reinforcement learning can greatly benefit from the use of options as a way of encoding recurring behaviours and to foster exploration. An important open problem is how can an agen…