Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Trust-SSL: Additive-Residual Selective Invariance for Robust Aerial Self-Supervised Learning
Wadii Boulila, Adel Ammar, Bilel Benjdira +1
Self-supervised learning (SSL) is a standard approach for representation learning in aerial imagery. Existing methods enforce invariance between augmented views, which works well w…
cs.CV2025
QARI-OCR: High-Fidelity Arabic Text Recognition through Multimodal Large Language Model Adaptation
Ahmed Wasfy, Omer Nacar, Abdelakreem Elkhateb +4
The inherent complexities of Arabic script; its cursive nature, diacritical marks (tashkeel), and varied typography, pose persistent challenges for Optical Character Recognition (O…
cs.CV2025
SARD: A Large-Scale Synthetic Arabic OCR Dataset for Book-Style Text Recognition
Omer Nacar, Yasser Al-Habashi, Serry Sibaee +2
Arabic Optical Character Recognition (OCR) is essential for converting vast amounts of Arabic print media into digital formats. However, training modern OCR models, especially powe…