benchmark dataset 1expert disagreement 1implicit reasoning 1legal citation detection 1model evaluation 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
Where Experts Disagree, Models Fail: Detecting Implicit Legal Citations in French Court Decisions
Avrile Floro, Tamara Dhorasoo, Soline Pellez +1
The paper introduces a benchmark for detecting implicit citations of the French Civil Code in court decisions and shows that cases where legal experts disagree are especially hard…
cs.AI2026
Cultural Binding Heads in Language Models
Avrile Floro, Luca Benedetto
LLMs often default to equal treatment across cultural groups, even though context warrants differentiation: this is a lack of difference awareness. Using mechanistic interpretabili…