1 paper · 1 filter
Kaiyuan Ji, Yijin Guo, Zicheng Zhang +4
With the increasing use of large language models (LLMs) in medical decision-support, it is essential to evaluate not only their final answers but also the reliability of their reas…