3 papers
cs.LG2026
On the Identifiability of Controlled World Models
Xiangteng Zhang, Yang Guan, Bo Zhang +3
World model serves as a promising tool to infer environment dynamics under high-dimensional observations and candidate actions. Recently, LeCun's JEPA provides a compelling framewo…
cs.LG2025
Distributional Soft Actor-Critic with Three Refinements
Jingliang Duan, Wenxuan Wang, Liming Xiao +6
Reinforcement learning (RL) has shown remarkable success in solving complex decision-making and control tasks. However, many model-free RL algorithms experience performance degrada…
cs.LG2025
Feasible Policy Iteration for Safe Reinforcement Learning
Yujie Yang, Zhilong Zheng, Shengbo Eben Li +4
Safety is the priority concern when applying reinforcement learning (RL) algorithms to real-world control problems. While policy iteration provides a fundamental algorithm for stan…