Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Improving Generalization Robustness of Multimodal RLVR
Pengfei Zhou, Zhiwei Tang, Xiaopeng Peng +11
Reinforcement Learning with Verifiable Rewards (RLVR) makes Multimodal Large Language Models more accurate, but the gains are brittle: simply paraphrasing a question or changing th…
cs.AI2025
EchoQA: A Large Collection of Instruction Tuning Data for Echocardiogram Reports
Lama Moukheiber, Mira Moukheiber, Dana Moukheiiber +2
We introduce a novel question-answering (QA) dataset using echocardiogram reports sourced from the Medical Information Mart for Intensive Care database. This dataset is specificall…