5 citations · 5 across the 3 of their papers we have counts for
3 papers
cs.LG2020
A Gentle Lecture Note on Filtrations in Reinforcement Learning
W. J. A. van Heeswijk
This note aims to provide a basic intuition on the concept of filtrations as used in the context of reinforcement learning (RL). Filtrations are often used to formally define RL pr…
cs.AI2020
Smart Containers With Bidding Capacity: A Policy Gradient Algorithm for Semi-Cooperative Learning
Wouter van Heeswijk
Smart modular freight containers -- as propagated in the Physical Internet paradigm -- are equipped with sensors, data storage capability and intelligence that enable them to route…
cs.LG2019★ 5 cited
Approximate Dynamic Programming with Neural Networks in Linear Discrete Action Spaces
Wouter van Heeswijk, Han La Poutré
Real-world problems of operations research are typically high-dimensional and combinatorial. Linear programs are generally used to formulate and efficiently solve these large decis…