2 citations · 2 across the 2 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024★ 2 cited
IdeaBench: Benchmarking Large Language Models for Research Idea Generation
Sikun Guo, Amir Hassan Shariatmadari, Guangzhi Xiong +4
Large Language Models (LLMs) have transformed how people interact with artificial intelligence (AI) systems, achieving state-of-the-art results in various tasks, including scientif…
cs.CL2022
'Tis but Thy Name: Semantic Question Answering Evaluation with 11M Names for 1M Entities
Albert Huang
Classic lexical-matching-based QA metrics are slowly being phased out because they punish succinct or informative outputs just because those answers were not provided as ground tru…