2 papers
cs.AI2026
LongDocBench: Benchmarking TOC Hierarchy and Contextual Relationship Recovery in Long Documents
Yuefeng Zou, Yichen Lu, Jingxiao Yang +5
Parsing visual documents into machine-readable representations is fundamental to document intelligence. Existing benchmarks focus on page-level element recognition, reading order,…
cs.CV2026
LingDT-VL-OCR: Structure-Aware Document-Level Parsing with Fine-Grained Visual Reference
Siyi Qian, Xiongfei Bai, Bingtao Fu +4
In this paper, we propose LingDT-VL-OCR, a document parsing system tailored to financial-domain documents, transforming ultra-long financial PDFs into semantically consistent, high…