2 citations · 5 across the 3 of their papers we have counts for
4 papers · 1 filter
LLM-Mini-CEX: Automatic Evaluation of Large Language Model for Diagnostic Conversation
Xiaoming Shi, Jie Xu, Jinru Ding +9
There is an increasing interest in developing LLMs for medical diagnosis to improve diagnosis efficiency. Despite their alluring technological potential, there is no unified and co…
MidMed: Towards Mixed-Type Dialogues for Medical Consultation
Xiaoming Shi, Zeming Liu, Chuan Wang +4
Most medical dialogue systems assume that patients have clear goals (medicine querying, surgical operation querying, etc.) before medical consultation. However, in many real scenar…
MedGPTEval: A Dataset and Benchmark to Evaluate Responses of Large Language Models in Medicine
Jie Xu, Lu Lu, Sen Yang +10
METHODS: First, a set of evaluation criteria is designed based on a comprehensive literature review. Second, existing candidate criteria are optimized for using a Delphi method by…
Language Models Can Improve Event Prediction by Few-Shot Abductive Reasoning
Xiaoming Shi, Siqiao Xue, Kangrui Wang +5
Large language models have shown astonishing performance on a wide range of reasoning tasks. In this paper, we investigate whether they could reason about real-world events and hel…