2 papers
cs.LG2026
A KL-regularization Framework for Learning to Plan with Adaptive Priors
Ãlvaro Serra-Gomez, Daniel Jarne Ornia, Dhruva Tirumala +1
Effective exploration remains a central challenge in model-based reinforcement learning (MBRL), particularly in high-dimensional continuous control tasks where sample efficiency is…
cs.AI2024
Reinforcement learning for Quantum Tiq-Taq-Toe
Catalin-Viorel Dinu, Thomas Moerland
Quantum Tiq-Taq-Toe is a well-known benchmark and playground for both quantum computing and machine learning. Despite its popularity, no reinforcement learning (RL) methods have be…