1 paper
Yi Yu, Bo Wang, Chong Feng +4
Current evaluations of large language models (LLMs) primarily focus on factual knowledge retrieval, overlooking the fundamental challenge of navigating the complex, non-bijective m…