2 papers
cs.IR2026
Projecting BrowseComp-Plus onto ClimbMix: Toward More Realistic Corpora for Agentic Search
Sahel Sharifymoghaddam, Lingwei Gu, Yijun Ge +1
The BrowseComp-Plus benchmark disentangled the evaluation of agentic search by replacing opaque web search with a fixed corpus, so that an agent's role can be separated from the re…
cs.IR2025
Lighting the Way for BRIGHT: Reproducible Baselines with Anserini, Pyserini, and RankLLM
Sahel Sharifymoghaddam, Yijun Ge, Jimmy Lin
Retrieval benchmarks for large language models (LLMs) should reflect the long, reasoning-intensive queries typical of retrieval-augmented generation (RAG). We present a systematic…