1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Patrick Vossler, Fan Xia, Yifan Mai +2
Given the challenge of automatically evaluating free-form outputs from large language models (LLMs), an increasingly common solution is to use LLMs themselves as the judging mechan…