2 papers
cs.CV2026
VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents
Udi Barzelay, Ophir Azulai, Inbar Shapira +4
We introduce VAREX (VARied-schema EXtraction), a benchmark for evaluating multimodal foundation models on structured data extraction from government forms. VAREX employs a Reverse…
cs.CL2024
Augmenting In-Context-Learning in LLMs via Automatic Data Labeling and Refinement
Joseph Shtok, Amit Alfassy, Foad Abo Dahood +3
It has been shown that Large Language Models' (LLMs) performance can be improved for many tasks using Chain of Thought (CoT) or In-Context Learning (ICL), which involve demonstrati…