Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
When (and How) to Trust the Expert: Diagnosing Query-Time Expert-Guided Reinforcement Learning
Yann Berthelot, Philippe Preux, Riad Akrour
Many continuous-control problems ship with a competent but suboptimal controller (a tuned PID, a hand-designed gait). A growing family of methods uses such controllers as queryable…
cs.AI2024
IDEQ -- Improving Diffusion Models for the Traveling Salesman Problem (TSP) by Leveraging the Structure of the Solution Space
Mickael Basson, Philippe Preux
We investigate diffusion models to solve the Traveling Salesman Problem. Building on the recent DIFUSCO and T2TCO approaches, we propose IDEQ. IDEQ improves the quality of the solu…
cs.AI2024
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
Hector Kohler, Quentin Delfosse, Riad Akrour +2
Deep reinforcement learning agents are prone to goal misalignments. The black-box nature of their policies hinders the detection and correction of such misalignments, and the trust…