bayesian methods 1hyperparameter tuning 1model-based RL 1offline reinforcement learning 1theoretical analysis 1
From the 1 of 8 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
When Do We Need LLMs? A Diagnostic for Language-Driven Bandits
Uljad Berdica, Fernando Acero, Anton Ipsen +3
We study Contextual Multi-Armed Bandits (CMABs) for non-episodic decision-making problems where the context includes both textual and numerical information (e.g., recommendation sy…
cs.AI2025
Intent Factored Generation: Unleashing the Diversity in Your Language Model
Eltayeb Ahmed, Uljad Berdica, Martha Elliott +2
Obtaining multiple meaningfully diverse, high quality samples from Large Language Models for a fixed prompt remains an open challenge. Current methods for increasing diversity ofte…