1 paper · 1 filter
Qian Zhang, Panfeng Chen, Jiali Li +6
The emergence of Large Language Models (LLMs) in the medical domain has stressed a compelling need for standard datasets to evaluate their question-answering (QA) performance. Alth…