◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

I. Gevers

2 papers hereh-index 13 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.CL2

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.CL2026

Benchmarking the Benchmarks: Testing the Predictive Validity of Commonsense Benchmarks

Ine Gevers, Walter Daelemans

Predicting LLM's capabilities on real-world tasks is essential, yet the extent to which performance on commonsense benchmarks predicts downstream performance remains underspecified…

cs.CL2026

Do You Get the Hint? Benchmarking LLMs on the Board Game Concept

Ine Gevers, Walter Daelemans

Large language models (LLMs) have achieved striking successes on many benchmarks, yet recent studies continue to expose fundamental weaknesses. In this paper, we introduce Concept,…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.