1 paper · 1 filter
Shouang Wei, Min Zhang, Xin Lin +3
Recently, several multi-turn dialogue benchmarks have been proposed to evaluate the conversational abilities of large language models (LLMs). As LLMs are increasingly recognized as…