2 papers
cs.LG2026
Command-Space Counterfactual Explanations for Pareto-Conditioned Reinforcement Learning
Joanikij Chulev, Hendrik Baier
Pareto Conditioned Networks learn multiple multi-objective reinforcement learning behaviours by conditioning a single policy on a desired return command. However, the local mapping…
cs.LG2025
InnateCoder: Learning Programmatic Options with Foundation Models
Rubens O. Moraes, Quazi Asif Sadmine, Hendrik Baier +1
Outside of transfer learning settings, reinforcement learning agents start their learning process from a clean slate. As a result, such agents have to go through a slow process to…