3 papers
cs.CL2026
When Text and Numbers Disagree: Evidence Arbitration in Large Language Models
Mattia Carletti, Edward Phillips, Fredrik K. Gustafsson +5
Large language models (LLMs) are increasingly used in settings where textual summaries, numerical observations, and external tool outputs may provide conflicting evidence. We study…
cs.LG2026
SignalMC-MED: A Multimodal Benchmark for Evaluating Biosignal Foundation Models on Single-Lead ECG and PPG
Fredrik K. Gustafsson, Xiao Gu, Mattia Carletti +3
Recent biosignal foundation models (FMs) have demonstrated promising performance across diverse clinical prediction tasks, yet systematic evaluation on long-duration multimodal dat…
cs.LG2025
RiskAgent: Synergizing Language Models with Validated Tools for Evidence-Based Risk Prediction
Fenglin Liu, Jinge Wu, Hongjian Zhou +9
Large Language Models (LLMs) achieve competitive results compared to human experts in medical examinations. However, it remains a challenge to apply LLMs to complex clinical decisi…