2 papers
cs.CL2025
Cross-Lingual SynthDocs: A Large-Scale Synthetic Corpus for Any to Arabic OCR and Document Understanding
Haneen Al-Homoud, Asma Ibrahim, Murtadha Al-Jubran +5
Cross-Lingual SynthDocs is a large-scale synthetic corpus designed to address the scarcity of Arabic resources for Optical Character Recognition (OCR) and Document Understanding (D…
cs.CL2025
Multi-Agent Interactive Question Generation Framework for Long Document Understanding
Kesen Wang, Daulet Toibazar, Abdulrahman Alfulayt +6
Document Understanding (DU) in long-contextual scenarios with complex layouts remains a significant challenge in vision-language research. Although Large Vision-Language Models (LV…