47 citations · 47 across the 4 of their papers we have counts for
4 papers
MARBLE: A Hard Benchmark for Multimodal Spatial Reasoning and Planning
Yulun Jiang, Yekun Chai, Maria Brbić +1
The ability to process information from multiple modalities and to reason through it step-by-step remains a critical challenge in advancing artificial intelligence. However, existi…
MIRIAD: Augmenting LLMs with millions of medical query-response pairs
Qinyue Zheng, Salman Abdullah, Sam Rawal +7
LLMs are bound to transform healthcare with advanced decision support and flexible chat assistants. However, LLMs are prone to generate inaccurate medical content. To ground LLMs i…
Style-Aware Radiology Report Generation with RadGraph and Few-Shot Prompting
Benjamin Yan, Ruochen Liu, David E. Kuo +8
Automatically generated reports from medical images promise to improve the workflow of radiologists. Existing methods consider an image-to-report modeling task by directly generati…
Med-Flamingo: a Multimodal Medical Few-shot Learner
Michael Moor, Qian Huang, Shirley Wu +6
Medicine, by its nature, is a multifaceted domain that requires the synthesis of information across various modalities. Medical generative vision-language models (VLMs) make a firs…