2 papers
stat.ML2026
E-Scores for (In)Correctness Assessment of Generative Model Outputs
Guneet S. Dhillon, Javier González, Teodora Pandeva +1
While generative models, especially large language models (LLMs), are ubiquitous in today's world, principled mechanisms to assess their (in)correctness are limited. Using the conf…
cs.LG2025
L3Ms -- Lagrange Large Language Models
Guneet S. Dhillon, Xingjian Shi, Yee Whye Teh +1
Supervised fine-tuning (SFT) and alignment of large language models (LLMs) are key steps in providing a good user experience. However, the concept of an appropriate alignment is in…