2 papers
cs.LG2026
Jump Start or False Start? A Theoretical and Empirical Evaluation of LLM-initialized Bandits
Adam Bayley, Xiaodan Zhu, Raquel Aoki +2
The recent advancement of Large Language Models (LLMs) offers new opportunities to generate user preference data to warm-start bandits. Recent studies on contextual bandits with LL…
cs.LG2025
No : Model-Agnostic Counterfactual Explanations Using Reinforcement Learning
Xiangyu Sun, Raquel Aoki, Kevin H. Wilson
Machine learning (ML) methods have experienced significant growth in the past decade, yet their practical application in high-impact real-world domains has been hindered by their o…