2 papers
cs.AI2026
LLMs for Medical Consultation Are Evaluated Too Late: The Preformulation Gap
Yining Hua, Cyrus Ayubcha, Hongbin Na +4
Large language models for medical consultation are often evaluated after a clinical problem has already been made clear, although real consultations may begin with a vague, minimiz…
cs.IR2025
A Multi-Agent Approach to Neurological Clinical Reasoning
Moran Sorka, Alon Gorenshtein, Dvir Aran +1
Large language models (LLMs) have shown promise in medical domains, but their ability to handle specialized neurological reasoning requires systematic evaluation. We developed a co…