1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2024★ 1 cited
Integration of cognitive tasks into artificial general intelligence test for large models
Youzhi Qu, Chen Wei, Penghui Du +10
During the evolution of large models, performance evaluation is necessarily performed to assess their capabilities and ensure safety before practical application. However, current…
cs.CL2023
Beyond Factuality: A Comprehensive Evaluation of Large Language Models as Knowledge Generators
Liang Chen, Yang Deng, Yatao Bian +4
Large language models (LLMs) outperform information retrieval techniques for downstream knowledge-intensive tasks when being prompted to generate world knowledge. However, communit…