2 papers
cs.CV2026
LingDT-VL-OCR: Structure-Aware Document-Level Parsing with Fine-Grained Visual Reference
Siyi Qian, Xiongfei Bai, Bingtao Fu +4
In this paper, we propose LingDT-VL-OCR, a document parsing system tailored to financial-domain documents, transforming ultra-long financial PDFs into semantically consistent, high…
cs.CV2026
Tracing Copied Pixels and Regularizing Patch Affinity in Copy Detection
Yichen Lu, Siwei Nie, Minlong Lu +3
Image Copy Detection (ICD) aims to identify manipulated content between image pairs through robust feature representation learning. While self-supervised learning (SSL) has advance…