1 paper
Robert J. Moore, Sungeun An, Farhan Ahmed +1
The Natural Conversation Benchmark (NC-Bench) introduces a new approach to evaluating the general conversational competence of large language models (LLMs). Unlike prior benchmarks…