18 papers
Explaining When PRF Fails: Participatory Auditing for Selective Query Expansion
Zeyan Liang, Graham McDonald, Iadh Ounis
The paper investigates why pseudo-relevance feedback (PRF) harms many queries and introduces a two‑stage audit‑then‑automate framework that uses user audits and LLM‑based rerankers…
Interpretable Uncertainty for Adaptive Retrieval and Reasoning in Question Answering
Ritajit Dey, Iadh Ounis, Graham McDonald
Large language models (LLMs) achieve a strong performance in question answering (QA), but remain prone to hallucinations and suffer from limited transparency. Retrieval-augmented g…
URecJPQ: Memory-efficient Multimodal Recommendation Models through RecJPQ in Large-Scale Scenarios
Giuseppe Spillo, Zixuan Yi, Aleksandr Petrov +3
Training state-of-the-art recommendation models on large-scale industrial datasets can be a challenging task due to the high number of users and items which are typically represent…
Certifiable Semantic Agreement Among LLM Agents: What the Admissibility Instrument Decides
Haoran Xu, Lei Zhang, Iadh Ounis +1
Can a committee of LLM agents reach agreement that is certifiable at the level of meaning, not only at the level of a label? We build a protocol to find out. H-CSC emits one of thr…
A Large-Scale Dataset and Benchmark: Do Protein-Ligand Models Learn Binding Sites or Just Binding Likelihood?
Zhaohan Meng, Zhen Bai, Ke Yuan +4
Protein-ligand modeling underpins computational drug discovery and molecular design. Existing protein-ligand benchmarks typically evaluate whether a protein and ligand interact and…
All Eyes on the Ranker: Participatory Auditing to Surface Blind Spots in Ranked Search Results
Anna Marie Rezk, Patrizia Di Campli San Vito, Ayah Soufan +3
Search engines that present users with a ranked list of search results are a fundamental technology for providing public access to information. Evaluations of such systems are typi…