2 papers
cs.LG2026
Lever: Inference-Time Policy Reuse under Support Constraints
Ihor Vitenko, Noha Ibrahim, Sihem Amer-Yahia
Reinforcement learning (RL) policies are typically trained for fixed objectives, making reuse difficult when task requirements change. We study inference-time policy reuse: given a…
cs.LG2026
Optimizing Coverage and Difficulty in Reinforcement Learning for Quiz Composition
Ricardo Pedro Querido Andrade Silva, Nassim Bouarour, Dina Fettache +3
Quiz design is a tedious process that teachers undertake to evaluate the acquisition of knowledge by students. Our goal in this paper is to automate quiz composition from a set of…