1 paper · 1 filter
Tianjun Wei, Wei Wen, Ruizhi Qiao +2
Evaluating large language models (LLMs) in diverse and challenging scenarios is essential to align them with human preferences. To mitigate the prohibitive costs associated with hu…