3 citations · 5 across the 4 of their papers we have counts for
4 papers
On the Structural Non-Preservation of Epistemic Behaviour under Policy Transformation
Alexander Galozy
Reinforcement learning (RL) agents under partial observability often condition actions on internally accumulated information such as memory or inferred latent context. We formalise…
Beyond Random Noise: Insights on Anonymization Strategies from a Latent Bandit Study
Alexander Galozy, Sadi Alawadi, Victor Kebande +1
This paper investigates the issue of privacy in a learning scenario where users share knowledge for a recommendation task. Our study contributes to the growing body of research on…
Information-Gathering in Latent Bandits
Alexander Galozy, Slawomir Nowaczyk
In the latent bandit problem, the learner has access to reward distributions and -- for the non-stationary variant -- transition models of the environment. The reward distributions…
A New Bandit Setting Balancing Information from State Evolution and Corrupted Context
Alexander Galozy, Slawomir Nowaczyk, Mattias Ohlsson
We propose a new sequential decision-making setting, combining key aspects of two established online learning problems with bandit feedback. The optimal action to play at any given…