2 citations · 4 across the 5 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Generalized Kalman filter based temporal difference reinforcement learning
Vasos Arnaoutis, Eric Lutters, Bojana Rosić
In this paper, we present a generalized temporal-difference (TD) reinforcement learning framework based on the theory of conditional expectations. The value and action-value (Q-val…
cs.LG2021★ 2 cited
Long short-term relevance learning
Bram van de Weg, Lars Greve, Bojana Rosic
To incorporate prior knowledge as well as measurement uncertainties in the traditional long short term memory (LSTM) neural networks, an efficient sparse Bayesian training algorith…