1 citations · 1 across the 1 of their papers we have counts for
1 paper
Antoni Lasik, Jakub Pokrywka, Łukasz Grzybowski +7
Large language models (LLMs) in medicine are mainly evaluated using multiple-choice question answering (MCQA), which can overestimate real clinical ability due to guessing strategi…