2 papers
cs.CL2025
A benchmark dataset for evaluating Syndrome Differentiation and Treatment in large language models
Kunning Li, Jianbin Guo, Zhaoyang Shang +5
The emergence of Large Language Models (LLMs) within the Traditional Chinese Medicine (TCM) domain presents an urgent need to assess their clinical application capabilities. Howeve…
cs.CL2025
TCM-3CEval: A Triaxial Benchmark for Assessing Responses from Large Language Models in Traditional Chinese Medicine
Tianai Huang, Lu Lu, Jiayuan Chen +5
Large language models (LLMs) excel in various NLP tasks and modern medicine, but their evaluation in traditional Chinese medicine (TCM) is underexplored. To address this, we introd…