4 papers
INFELM: In-depth Fairness Evaluation of Large Text-To-Image Models
Di Jin, Xing Liu, Yu Liu +7
The rapid development of large language models (LLMs) and large vision models (LVMs) have propelled the evolution of multi-modal AI systems, which have demonstrated the remarkable…
WorldMedQA-V: a multilingual, multimodal medical examination dataset for multimodal language models evaluation
João Matos, Shan Chen, Siena Placino +13
Multimodal/vision language models (VLMs) are increasingly being deployed in healthcare settings worldwide, necessitating robust benchmarks to ensure their safety, efficacy, and fai…
Evaluating the Impact of Pulse Oximetry Bias in Machine Learning under Counterfactual Thinking
Inês Martins, João Matos, Tiago Gonçalves +3
Algorithmic bias in healthcare mirrors existing data biases. However, the factors driving unfairness are not always known. Medical devices capture significant amounts of data but a…
EHRmonize: A Framework for Medical Concept Abstraction from Electronic Health Records using Large Language Models
João Matos, Jack Gallifant, Jian Pei +1
Electronic health records (EHRs) contain vast amounts of complex data, but harmonizing and processing this information remains a challenging and costly task requiring significant c…