3 papers
cs.CL2025
HealthContradict: Evaluating Biomedical Knowledge Conflicts in Language Models
Boya Zhang, Alban Bornet, Rui Yang +2
How do language models use contextual information to answer health questions? How are their responses impacted by conflicting contexts? We assess the ability of language models to…
cs.LG2025
ICU-TSB: A Benchmark for Temporal Patient Representation Learning for Unsupervised Stratification into Patient Cohorts
Dimitrios Proios, Alban Bornet, Anthony Yazdani +2
Patient stratification identifying clinically meaningful subgroups is essential for advancing personalized medicine through improved diagnostics and treatment strategies. Electroni…
cs.CL2025
An Evaluation Benchmark for Adverse Drug Event Prediction from Clinical Trial Results
Anthony Yazdani, Alban Bornet, Philipp Khlebnikov +4
Adverse drug events (ADEs) are a major safety issue in clinical trials. Thus, predicting ADEs is key to developing safer medications and enhancing patient outcomes. To support this…