2 papers
cs.LG2025
Impatient Bandits: Optimizing for the Long-Term Without Delay
Kelly W. Zhang, Thomas Baldwin-McDonald, Kamil Ciosek +2
Increasingly, recommender systems are tasked with improving users' long-term satisfaction. In this context, we study a content exploration task, which we formalize as a bandit prob…
cs.LG2024
On the Importance of Uncertainty in Decision-Making with Large Language Models
Nicolò Felicioni, Lucas Maystre, Sina Ghiassian +1
We investigate the role of uncertainty in decision-making problems with natural language as input. For such tasks, using Large Language Models as agents has become the norm. Howeve…