226 citations · 236 across the 7 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2024
Stitching Sub-Trajectories with Conditional Diffusion Model for Goal-Conditioned Offline RL
Sungyoon Kim, Yunseon Choi, Daiki E. Matsunaga +1
Offline Goal-Conditioned Reinforcement Learning (Offline GCRL) is an important problem in RL that focuses on acquiring diverse goal-oriented skills solely from pre-collected behavi…
cs.AI2014★ 226 cited
Learning to Cooperate via Policy Search
Leonid Peshkin, Kee-Eung Kim, Nicolas Meuleau +1
Cooperative games are those in which both agents share the same payoff structure. Value-based reinforcement-learning algorithms, such as variants of Q-learning, have been applied t…
cs.AI2012★ 2 cited
A Geometric Traversal Algorithm for Reward-Uncertain MDPs
Eunsoo Oh, Kee-Eung Kim
Markov decision processes (MDPs) are widely used in modeling decision making problems in stochastic environments. However, precise specification of the reward functions in MDPs is…