Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Trust Region Reward Optimization and Proximal Inverse Reward Optimization Algorithm
Yang Chen, Menglin Zou, Jiaqi Zhang +6
Inverse Reinforcement Learning (IRL) learns a reward function to explain expert demonstrations. Modern IRL methods often use the adversarial (minimax) formulation that alternates b…
cs.LG2024
Robust Domain Generalisation with Causal Invariant Bayesian Neural Networks
Gaël Gendron, Michael Witbrock, Gillian Dobbie
Deep neural networks can obtain impressive performance on various tasks under the assumption that their training domain is identical to their target domain. Performance can drop dr…