1 paper · 1 filter
Jui-Ming Yao, Hao-Yuan Chen, Zi-Xian Tang +4
Large Language Models (LLMs) have demonstrated impressive performance on multiple-choice question answering (MCQA) benchmarks, yet they remain highly vulnerable to minor input pert…