3 papers
cs.CV2026
Stringalign: Moving beyond summary statistics with a transparent Unicode-aware tool for evaluating automatic transcription models
Yngve Mardal Moe, Marie Roald
Comparing text strings is crucial when evaluating and understanding the performance of various text processing tasks such as document recognition and audio transcription. With an i…
cs.CL2025
Comparative analysis of optical character recognition methods for Sámi texts from the National Library of Norway
Tita Enstad, Trond Trosterud, Marie Iversdatter Røsok +2
Optical Character Recognition (OCR) is crucial to the National Library of Norway's (NLN) digitisation process as it converts scanned documents into machine-readable text. However,…
cs.CV2024
Visual Navigation of Digital Libraries: Retrieval and Classification of Images in the National Library of Norway's Digitised Book Collection
Marie Roald, Magnus Breder Birkenes, Lars Gunnarsønn Bagøien Johnsen
Digital tools for text analysis have long been essential for the searchability and accessibility of digitised library collections. Recent computer vision advances have introduced s…