Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
PRGB Benchmark: A Robust Placeholder-Assisted Algorithm for Benchmarking Retrieval-Augmented Generation
Zhehao Tan, Yihan Jiao, Dan Yang +7
Retrieval-Augmented Generation (RAG) enhances large language models (LLMs) by integrating external knowledge, where the LLM's ability to generate responses based on the combination…
cs.CL2024
A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions
Lei Liu, Xiaoyan Yang, Junchi Lei +6
With the advent of Large Language Models (LLMs), medical artificial intelligence (AI) has experienced substantial technological progress and paradigm shifts, highlighting the poten…
cs.CL2024
Towards Automatic Evaluation for LLMs' Clinical Capabilities: Metric, Data, and Algorithm
Lei Liu, Xiaoyan Yang, Fangzhou Li +10
Large language models (LLMs) are gaining increasing interests to improve clinical efficiency for medical diagnosis, owing to their unprecedented performance in modelling natural la…