3 citations · 4 across the 4 of their papers we have counts for
1 paper · 1 filter
Yuqi Wang, Boran Jiang, Yi Luo +3
Large language models (LLMs), such as GPT3.5, GPT4 and LLAMA2 perform surprisingly well and outperform human experts on many tasks. However, in many domain-specific evaluations, th…