346 citations
- Georgia Institute of TechnologyUS35 papers
- University of PennsylvaniaUS11 papers
- Atlanta University CenterUS10 papers
- Yale UniversityUS9 papers
- California Institute of TechnologyUS7 papers
- M.N. Mikheev Institute of Metal PhysicsRU7 papers
- Princeton UniversityUS7 papers
- University of ChicagoUS7 papers
- University of Illinois Urbana-ChampaignUS7 papers
- Bar-Ilan UniversityIL6 papers
- CeNTechDE6 papers
- Duke UniversityUS6 papers
Showing 2025 · cs.CLShow all
2 papers · 2 filters
cs.CL2025
PeruMedQA: Benchmarking Large Language Models (LLMs) on Peruvian Medical Exams -- Dataset Construction and Evaluation
Rodrigo M. Carrillo-Larco, Jesus Lovón Melgarejo, Manuel Castillo-Cara +1
BACKGROUND: Medical large language models (LLMs) have demonstrated remarkable performance in answering medical examinations. However, the extent to which this high performance is t…
cs.CL2025★ 14 cited
Measuring Sycophancy of Language Models in Multi-turn Dialogues
Jiseung Hong, Grace Byun, Seungone Kim +2
Large Language Models (LLMs) are expected to provide helpful and harmless responses, yet they often exhibit sycophancy--conforming to user beliefs regardless of factual accuracy or…