1 paper
Qi Cheng, Michael Boratko, Pranay Kumar Yelugam +4
Large language models have demonstrated impressive performance on commonsense tasks; however, these tasks are often posed as multiple-choice questions, allowing models to exploit s…