7 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.CL2022★ 1 cited
Radically Lower Data-Labeling Costs for Visually Rich Document Extraction Models
Yichao Zhou, James B. Wendt, Navneet Potti +2
A key bottleneck in building automatic extraction models for visually rich documents like invoices is the cost of acquiring the several thousand high-quality labeled documents that…
cs.LG2020★ 7 cited
Active Learning for Skewed Data Sets
Abbas Kazerouni, Qi Zhao, Jing Xie +2
Consider a sequential active learning problem where, at each round, an agent selects a batch of unlabeled data points, queries their labels and updates a binary classifier. While t…