4 papers
HiPS: Hierarchical PDF Segmentation of Doctrinal Legal Books
Sabine Wehnert, Harikrishnan Changaramkulath, Ivan Habernal
PDF parsers have recently improved on page-level layout understanding. However, recovering a document-global section hierarchy with reliable boundaries remains brittle for deeply s…
GDPR Auto-Formalization with AI Agents and Human Verification
Ha Thanh Nguyen, Wachara Fungwacharakorn, Sabine Wehnert +6
We study the overall process of automatic formalization of GDPR provisions using large language models, within a human-in-the-loop verification framework. Rather than aiming for fu…
Can Legislation Be Made Machine-Readable in PROLEG?
May-Myo Zin, Sabine Wehnert, Yuntao Kong +7
The anticipated positive social impact of regulatory processes requires both the accuracy and efficiency of their application. Modern artificial intelligence technologies, includin…
Analyzing Bias in Swiss Federal Supreme Court Judgments Using Facebook's Holistic Bias Dataset: Implications for Language Model Training
Sabine Wehnert, Muhammet Ertas, Ernesto William De Luca
Natural Language Processing (NLP) is vital for computers to process and respond accurately to human language. However, biases in training data can introduce unfairness, especially…