2 papers
cs.CL2025
LLM-as-a-Judge is Bad, Based on AI Attempting the Exam Qualifying for the Member of the Polish National Board of Appeal
Michał Karp, Anna Kubaszewska, Magdalena Król +4
This study provides an empirical assessment of whether current large language models (LLMs) can pass the official qualifying examination for membership in Poland's National Appeal…
cs.CL2025
Are manual annotations necessary for statutory interpretations retrieval?
Aleksander Smywiński-Pohl, Tomer Libal, Adam Kaczmarczyk +1
One of the elements of legal research is looking for cases where judges have extended the meaning of a legal concept by providing interpretations of what a concept means or does no…