large language models 2accountability 1agent-mediated collaboration 1contribution attribution 1evaluation framework 1knowledge work 1medical dialogue 1misconception detection 1multi-turn conversation 1
From the 2 of 16 linked papers with an AI index.
2 citations · 2 across the 9 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Evaluating Large Language Models on Misconceptions in Multi-Turn Medical Conversations
Monica Munnangi, Saiph Savage
The paper introduces ThReadMed-QA, a multi‑turn medical dialogue dataset, and evaluates how well large language models can detect and correct patient misconceptions across conversa…
cs.CL2026
ThReadMed-QA: A Multi-Turn Medical Dialogue Benchmark from Real Patient Questions
Monica Munnangi, Saiph Savage
Medical question-answering benchmarks predominantly evaluate single-turn exchanges, failing to capture the iterative, clarification-seeking nature of real patient consultations. We…