9 citations · 13 across the 4 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.CL2024
BuDDIE: A Business Document Dataset for Multi-task Information Extraction
Ran Zmigrod, Dongsheng Wang, Mathieu Sibue +10
The field of visually rich document understanding (VRDU) aims to solve a multitude of well-researched NLP tasks in a multi-modal domain. Several datasets exist for research on spec…
cs.CL2024★ 4 cited
Large Language Models as Financial Data Annotators: A Study on Effectiveness and Efficiency
Toyin Aguda, Suchetha Siddagangappa, Elena Kochkina +4
Collecting labeled datasets in finance is challenging due to scarcity of domain experts and higher cost of employing them. While Large Language Models (LLMs) have demonstrated rema…