2 citations · 6 across the 45 of their papers we have counts for
1 paper · 2 filters
Silvia Cappelletti, Tobia Poppi, Samuele Poppi +5
Large Language Models (LLMs) are increasingly evaluated on multiple-choice question answering (MCQA) tasks using *first-token probability* (FTP), which selects the answer option wh…