Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Is Q-learning an Ill-posed Problem?
Philipp Wissmann, Daniel Hein, Steffen Udluft +1
This paper investigates the instability of Q-learning in continuous environments, a challenge frequently encountered by practitioners. Traditionally, this instability is attributed…
cs.LG2024
Why long model-based rollouts are no reason for bad Q-value estimates
Philipp Wissmann, Daniel Hein, Steffen Udluft +1
This paper explores the use of model-based offline reinforcement learning with long model rollouts. While some literature criticizes this approach due to compounding errors, many p…