4 papers
Accelerometry-Derived Digital Biomarkers for Cardiometabolic Risk: A Population-Representative Tabular Benchmark with Uncertainty Quantification
Federico Felizzi
Structured tabular data dominates clinical medicine, yet existing benchmarks fail to reflect real-world properties like complex survey sampling, demographic oversampling, and subgr…
EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluation
Francesco Andrea Causio, Vittorio De Vita, Olivia Riccomi +17
While Large Language Models (LLMs) have demonstrated high proficiency on English-centric medical examinations, their performance often declines when faced with non-English language…
Are Large Vision Language Models Truly Grounded in Medical Images? Evidence from Italian Clinical Visual Question Answering
Federico Felizzi, Olivia Riccomi, Michele Ferramola +12
Large vision language models (VLMs) have achieved impressive performance on medical visual question answering benchmarks, yet their reliance on visual information remains unclear.…
Economic impact of biomarker-based aging interventions on healthcare costs and individual value
Federico Felizzi
We investigate the economic impact of controlling the pace of aging through biomarker monitoring and targeted interventions. Using the DunedinPACE epigenetic clock as a measure of…