collaborators
Showing cs.IRShow all

11 papers · 1 filter

cs.IR2026

Auto-ARGUE: LLM-Based Report Generation Evaluation

William Walden, Marc Mason, Orion Weller +10

Generation of citation-backed reports is a primary use case for retrieval-augmented generation (RAG) systems. While open-source evaluation tools exist for various RAG tasks, tools…

cs.IR2026

Milco: Learned Sparse Retrieval Across Languages via a Multilingual Connector

Thong Nguyen, Yibin Lei, Jia-Huei Ju +2

Learned Sparse Retrieval (LSR) combines the efficiency of bi-encoders with the transparency of lexical matching, but existing approaches struggle to scale beyond English. We introd…

cs.IR2026

Topic-Specific Classifiers are Better Relevance Judges than Prompted LLMs

Lukas Gienapp, Martin Potthast, Andrew Yates +2

The unjudged document problem, where systems that did not contribute to the original judgement pool may retrieve documents without a relevance judgement, is a key obstacle to the r…

cs.IR2025

Augmenting Researchy Questions with Sub-question Judgments

Jia-Huei Ju, Eugene Yang, Trevor Adriaanse +1

The Researchy Questions dataset provides about 100k question queries with complex information needs that require retrieving information about several aspects of a topic. Each query…

cs.IR2025

Overview of the TREC 2024 NeuCLIR Track

Dawn Lawrie, Sean MacAvaney, James Mayfield +4

The principal goal of the TREC Neural Cross-Language Information Retrieval (NeuCLIR) track is to study the effect of neural approaches on cross-language information access. The tra…

cs.IR2025

Rank1: Test-Time Compute for Reranking in Information Retrieval

Orion Weller, Kathryn Ricci, Eugene Yang +3

We introduce Rank1, the first reranking model trained to take advantage of test-time compute. Rank1 demonstrates the applicability within retrieval of using a reasoning language mo…