5 papers
The FACTS Leaderboard: A Comprehensive Benchmark for Large Language Model Factuality
Aileen Cheng, Alon Jacovi, Amir Globerson +62
We introduce The FACTS Leaderboard, an online leaderboard suite and associated set of benchmarks that comprehensively evaluates the ability of language models to generate factually…
Harnessing Pairwise Ranking Prompting Through Sample-Efficient Ranking Distillation
Junru Wu, Le Yan, Zhen Qin +6
While Pairwise Ranking Prompting (PRP) with Large Language Models (LLMs) is one of the most effective zero-shot document ranking methods, it has a quadratic computational complexit…
Optimizing Compound Retrieval Systems
Harrie Oosterhuis, Rolf Jagerman, Zhen Qin +1
Modern retrieval systems do not rely on a single ranking model to construct their rankings. Instead, they generally take a cascading approach where a sequence of ranking models are…
Adapting Decoder-Based Language Models for Diverse Encoder Downstream Tasks
Paul Suganthan, Fedor Moiseev, Le Yan +7
Decoder-based transformers, while revolutionizing language modeling and scaling to immense sizes, have not completely overtaken encoder-heavy architectures in natural language proc…
Searching Personal Collections
Michael Bendersky, Donald Metzler, Marc Najork +1
This article describes the history of information retrieval on personal document collections.