8 citations · 8 across the 4 of their papers we have counts for
4 papers
Augmented Hypothesis Testing with Persona-Based LLM Simulations
Ziyad Benomar, Aymen Al Marjani, Paul Missault +1
A/B testing requires large sample sizes, long timelines, and significant costs. When auxiliary predictions of experimental outcomes are available from machine learning models, unce…
Optimistic PAC Reinforcement Learning: the Instance-Dependent View
Andrea Tirinzoni, Aymen Al-Marjani, Emilie Kaufmann
Optimistic algorithms have been extensively studied for regret minimization in episodic tabular MDPs, both from a minimax and an instance-dependent view. However, for the PAC RL pr…
Near Instance-Optimal PAC Reinforcement Learning for Deterministic MDPs
Andrea Tirinzoni, Aymen Al-Marjani, Emilie Kaufmann
In probably approximately correct (PAC) reinforcement learning (RL), an agent is required to identify an -optimal policy with probability . While minimax optimal algorithms…
Adaptive Sampling for Best Policy Identification in Markov Decision Processes
Aymen Al Marjani, Alexandre Proutiere
We investigate the problem of best-policy identification in discounted Markov Decision Processes (MDPs) when the learner has access to a generative model. The objective is to devis…