2 papers
cs.CL2024
"What is the value of {templates}?" Rethinking Document Information Extraction Datasets for LLMs
Ran Zmigrod, Pranav Shetty, Mathieu Sibue +4
The rise of large language models (LLMs) for visually rich document understanding (VRDU) has kindled a need for prompt-response, document-based datasets. As annotating new datasets…
cs.AI2024
Fine-Tuning Language Models with Differential Privacy through Adaptive Noise Allocation
Xianzhi Li, Ran Zmigrod, Zhiqiang Ma +2
Language models are capable of memorizing detailed patterns and information, leading to a double-edged effect: they achieve impressive modeling performance on downstream tasks with…