Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models
Boqi Chen, Xudong Liu, Jiachuan Peng +5
Multimodal large language models (MLLMs) have shown great potential in medical applications, yet existing benchmarks inadequately capture real-world clinical complexity. We introdu…
cs.CL2025
Energy Landscapes Enable Reliable Abstention in Retrieval-Augmented Large Language Models for Healthcare
Ravi Shankar, Sheng Wong, Lin Li +4
Reliable abstention is critical for retrieval-augmented generation (RAG) systems, particularly in safety-critical domains such as women's health, where incorrect answers can lead t…