2 papers
cs.CL2023
LM-Polygraph: Uncertainty Estimation for Language Models
Ekaterina Fadeeva, Roman Vashurin, Akim Tsvigun +9
Recent advancements in the capabilities of large language models (LLMs) have paved the way for a myriad of groundbreaking applications in various fields. However, a significant cha…
cs.LG2023
One-Step Distributional Reinforcement Learning
Mastane Achab, Reda Alami, Yasser Abdelaziz Dahou Djilali +2
Reinforcement learning (RL) allows an agent interacting sequentially with an environment to maximize its long-term expected return. In the distributional RL (DistrRL) paradigm, the…