2 papers
cs.CL2026
CARE: A Conformal Safety Layer for Medical Summarization
Suhana Bedi, Bridget Lin, Anson Y. Zhou +5
Large language models (LLMs) are increasingly used for medical summarization, but their outputs can omit medically important information and introduce unsupported claims. Existing…
cs.AI2025
SycEval: Evaluating LLM Sycophancy
Aaron Fanous, Jacob Goldberg, Ank A. Agarwal +4
Large language models (LLMs) are increasingly applied in educational, clinical, and professional settings, but their tendency for sycophancy -- prioritizing user agreement over ind…