2 papers
cs.CL2026
CliniBench: A Clinical Outcome Prediction Benchmark for Generative and Encoder-Based Language Models
Paul Grundmann, Dennis Fast, Jan Frick +4
With their growing capabilities, generative large language models (LLMs) are being increasingly investigated for complex medical tasks. However, their effectiveness in real-world c…
cs.CL2025
MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data
Tianyu Han, Lisa C. Adams, Jens-Michalis Papaioannou +6
As large language models (LLMs) like OpenAI's GPT series continue to make strides, we witness the emergence of artificial intelligence applications in an ever-expanding range of fi…