1 paper · 1 filter
Seungpil Lee, Woochang Sim, Donghyeon Shin +6
The existing methods for evaluating the inference abilities of Large Language Models (LLMs) have been predominantly results-centric, making it challenging to assess the inference p…