1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Do Xuan Long, Hai Nguyen Ngoc, Tiviatis Sim +5
We present the first systematic evaluation examining format bias in performance of large language models (LLMs). Our approach distinguishes between two categories of an evaluation…