2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.LG2024★ 1 cited
Reinforcement Learning in Categorical Cybernetics
Jules Hedges, Riu Rodríguez Sakamoto
We show that several major algorithms of reinforcement learning (RL) fit into the framework of categorical cybernetics, that is to say, parametrised bidirectional processes. We bui…
math.CT2022★ 2 cited
Value Iteration is Optic Composition
Jules Hedges, Riu Rodríguez Sakamoto
Dynamic programming is a class of algorithms used to compute optimal control policies for Markov decision processes. Dynamic programming is ubiquitous in control theory, and is als…