6 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CL2025
LiveSearchBench: An Automatically Constructed Benchmark for Retrieval and Reasoning over Dynamic Knowledge
Heng Zhou, Ao Yu, Yuchen Fan +10
Evaluating large language models (LLMs) on question answering often relies on static benchmarks that reward memorization and understate the role of retrieval, failing to capture th…
cs.LG2022★ 6 cited
MineRL Diamond 2021 Competition: Overview, Results, and Lessons Learned
Anssi Kanervisto, Stephanie Milani, Karolis Ramanauskas +19
Reinforcement learning competitions advance the field by providing appropriate scope and support to develop solutions toward a specific problem. To promote the development of more…