2 papers
cs.LG2026
GEM: Guided Expectation-Maximization for Behavior-Normalized Candidate Action Selection in Offline RL
Haoyu Wang, Jingcheng Wang, Shunyu Wu +1
Offline reinforcement learning (RL) can fit strong value functions from fixed datasets, yet reliable deployment still hinges on the action selection interface used to query them. W…
cs.RO2026
DRIFT: Diffusion-based Rule-Inferred For Trajectories
Jinyang Zhao, Handong Zheng, Yanjiu Zhong +3
Trajectory generation for mobile robots in unstructured environments faces a critical dilemma: balancing kinematic smoothness for safe execution with terminal precision for fine-gr…