Showing cs.AIShow all
2 papers · 1 filter
cs.AI2003
Temporal plannability by variance of the episode length
Balint Takacs, Istvan Szita, Andras Lorincz
Optimization of decision problems in stochastic environments is usually concerned with maximizing the probability of achieving the goal and minimizing the expected episode length.…
cs.AI2002
Searching for Plannable Domains can Speed up Reinforcement Learning
Istvan Szita, Balint Takacs, Andras Lorincz
Reinforcement learning (RL) involves sequential decision making in uncertain environments. The aim of the decision-making agent is to maximize the benefit of acting in its environm…