equivariance 1flow matching 1generative models 1markov decision process 1motion planning 1multi-robot coordination 1persistent monitoring 1reinforcement learning 1robot manipulation 1se(2 1weighted latency 1
From the 2 of 11 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Analysis of On-policy Policy Gradient Methods under the Distribution Mismatch
Weizhen Wang, Jianping He, Xiaoming Duan
Policy gradient methods are one of the most successful approaches for solving challenging reinforcement learning problems. Despite their empirical successes, many state-of-the-art…
cs.LG2025
Model Selection for Inverse Reinforcement Learning via Structural Risk Minimization
Chendi Qu, Jianping He, Xiaoming Duan +1
Inverse reinforcement learning (IRL) usually assumes the reward function model is pre-specified as a weighted sum of features and estimates the weighting parameters only. However,…