55 citations · 79 across the 3 of their papers we have counts for
3 papers
cs.CL2026
QuarkMedBench: A Real-World Scenario Driven Benchmark for Evaluating Large Language Models
Yao Wu, Kangping Yin, Liang Dong +13
While Large Language Models (LLMs) excel on standardized medical exams, high scores often fail to translate to high-quality responses for real-world medical queries. Current evalua…
cs.CL2021★ 24 cited
CBLUE: A Chinese Biomedical Language Understanding Evaluation Benchmark
Ningyu Zhang, Mosha Chen, Zhen Bi +20
Artificial Intelligence (AI), along with the recent progress in biomedical language understanding, is gradually changing medical practice. With the development of biomedical langua…
cs.CL2020★ 55 cited
Conceptualized Representation Learning for Chinese Biomedical Text Mining
Ningyu Zhang, Qianghuai Jia, Kangping Yin +3
Biomedical text mining is becoming increasingly important as the number of biomedical documents and web data rapidly grows. Recently, word representation models such as BERT has ga…