7 papers
Vibe coding for clinicians: democratising bespoke software development for digital health innovation
Ariel Yuhan Ong, Iain Livingstone, Caroline Kilduff +5
Clinicians often face workflow problems that are perceived as either too bespoke or low stakes to attract commercial attention. Historically, most do not have the technical knowled…
Deliberative multi-agent large language models improve clinical reasoning in ophthalmology
Ehsan Misaghi, Sean T Berkowitz, Bing Yu Chen +10
Large language models (LLMs) show potential for ophthalmic clinical reasoning, yet individual models risk introducing harm. We evaluated whether multi-agent LLM deliberative counci…
Complementary Human-AI Clinical Reasoning in Ophthalmology
Mertcan Sevgi, Fares Antaki, Abdullah Zafar Khan +26
Vision impairment and blindness are a major global health challenge where gaps in the ophthalmology workforce limit access to specialist care. We evaluate AMIE, a medically fine-tu…
Performance of GPT-5 Frontier Models in Ophthalmology Question Answering
Fares Antaki, David Mikhail, Daniel Milad +11
Large language models (LLMs) such as GPT-5 integrate advanced reasoning capabilities that may improve performance on complex medical question-answering tasks. For this latest gener…
Benchmarking Next-Generation Reasoning-Focused Large Language Models in Ophthalmology: A Head-to-Head Evaluation on 5,888 Items
Minjie Zou, Sahana Srinivasan, Thaddaeus Wai Soon Lo +13
Recent advances in reasoning-focused large language models (LLMs) mark a shift from general LLMs toward models designed for complex decision-making, a crucial aspect in medicine. H…
Large language models perpetuate bias in palliative care: development and analysis of the Palliative Care Adversarial Dataset (PCAD)
Naomi Akhras, Fares Antaki, Fannie Mottet +4
Bias and inequity in palliative care disproportionately affect marginalised groups. Large language models (LLMs), such as GPT-4o, hold potential to enhance care but risk perpetuati…