17 citations · 28 across the 5 of their papers we have counts for
3 papers · 1 filter
Structured Q-learning For Antibody Design
Alexander I. Cowen-Rivers, Philip John Gorinski, Aivar Sootla +5
Optimizing combinatorial structures is core to many real-world problems, such as those encountered in life sciences. For example, one of the crucial steps involved in antibody desi…
Reinforcement Learning in Presence of Discrete Markovian Context Evolution
Hang Ren, Aivar Sootla, Taher Jafferjee +3
We consider a context-dependent Reinforcement Learning (RL) setting, which is characterized by: a) an unknown finite number of not directly observable contexts; b) abrupt (disconti…
SAMBA: Safe Model-Based & Active Reinforcement Learning
Alexander I. Cowen-Rivers, Daniel Palenicek, Vincent Moens +4
In this paper, we propose SAMBA, a novel framework for safe reinforcement learning that combines aspects from probabilistic modelling, information theory, and statistics. Our metho…