4 papers
Persona Conditioning as an Assessor-Sensitivity Probe for LLM-Based IR Evaluation
Samaneh Mohtadi, Pietro Bernardelle, Joel Mackenzie +1
Large language models (LLMs) are increasingly used as relevance assessors in information retrieval (IR) evaluation, raising questions about how assessor framing affects judgment re…
LLMs Encode Relevance as a Layer-Wise Cross-Lingual Signal
Pietro Bernardelle, Samaneh Mohtadi, Stefano Civelli +2
Large language models (LLMs) are increasingly used in information retrieval (IR) pipelines as relevance judges and re-rankers. Yet most analyses remain output-centric, evaluating g…
Query-Document Dense Vectors for LLM Relevance Judgment Bias Analysis
Samaneh Mohtadi, Gianluca Demartini
Large Language Models (LLMs) have been used as relevance assessors for Information Retrieval (IR) evaluation collection creation due to reduced cost and increased scalability as co…
The Effect of Document Summarization on LLM-Based Relevance Judgments
Samaneh Mohtadi, Kevin Roitero, Stefano Mizzaro +1
Relevance judgments are central to the evaluation of Information Retrieval (IR) systems, but obtaining them from human annotators is costly and time-consuming. Large Language Model…