2 papers
cs.CL2026
Post-training Large Language Models for Diverse High-Quality Responses
Yilei Chen, Souradip Chakraborty, Lorenz Wolf +2
Reinforcement learning (RL) has emerged as a popular method for post-training large language models (LLMs). While improving the model's performance on downstream tasks, it often re…
cs.CR2025
Private Selection with Heterogeneous Sensitivities
Daniela Antonova, Allegra Laro, Audra McMillan +1
Differentially private (DP) selection involves choosing a high-scoring candidate from a finite candidate pool, where each score depends on a sensitive dataset. This problem arises…