Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Self-Reported Confidence of Large Language Models in Gastroenterology: Analysis of Commercial, Open-Source, and Quantized Models
Nariman Naderi, Seyed Amir Ahmad Safavi-Naini, Thomas Savage +4
This study evaluated self-reported response certainty across several large language models (GPT, Claude, Llama, Phi, Mistral, Gemini, Gemma, and Qwen) using 300 gastroenterology bo…
cs.CL2024
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models
Seyed Amir Ahmad Safavi-Naini, Shuhaib Ali, Omer Shahab +15
Background and Aims: This study evaluates the medical reasoning performance of large language models (LLMs) and vision language models (VLMs) in gastroenterology. Methods: We used…