2 papers
cs.CL2026
AIPatient Arena: EHR-grounded evaluation of large language models in end-to-end clinical consultation workflows
Jiahui Niu, Huizi Yu, Wenkong Wang +11
Large language models (LLMs) are increasingly considered for use in clinical consultation tasks, yet most medical evaluations remain static, single-turn, or narrowly outcome-based,…
cs.CL2026
An evidence-guided reinforcement learning method to improve psychiatric reasoning in small language models
Xinxin Lin, Guangxin Dai, Yi Zhong +25
Privacy and computational constraints limit the use of large language models in psychiatry, while adapting small language models (SLMs) often requires substantial data and expert a…