1 paper
Sina Mansouri, Mohit Marvania, Vibhavari Ashok Shihorkar +5
Medical large language models (LLMs) achieve impressive performance on standardized benchmarks, yet these evaluations fail to capture the complexity of real clinical encounters whe…