Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation
Runze Zhao, Yue Yu, Adams Yiyue Zhu +2
Continuous-time reinforcement learning (CTRL) provides a principled framework for sequential decision-making in environments where interactions evolve continuously over time. Despi…
cs.LG2025
Provable Zero-Shot Generalization in Offline Reinforcement Learning
Zhiyong Wang, Chen Yang, John C. S. Lui +1
In this work, we study offline reinforcement learning (RL) with zero-shot generalization property (ZSG), where the agent has access to an offline dataset including experiences from…
cs.LG2024
CoPS: Empowering LLM Agents with Provable Cross-Task Experience Sharing
Chen Yang, Chenyang Zhao, Quanquan Gu +1
Sequential reasoning in agent systems has been significantly advanced by large language models (LLMs), yet existing approaches face limitations. Reflection-driven reasoning relies…