2 citations · 2 across the 3 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Beyond Accuracy: Benchmarking Cross-Task Consistency in Unified Multimodal Models
Weixing Wang, Liudvikas Zekas, Anton Hackl +5
Unified Multimodal Models (uMMs) aim to support both visual understanding and visual generation within a shared representation. However, existing evaluation protocols assess these…
cs.CV2023
DocLangID: Improving Few-Shot Training to Identify the Language of Historical Documents
Furkan Simsek, Brian Pfitzmann, Hendrik Raetz +3
Language identification describes the task of recognizing the language of written text in documents. This information is crucial because it can be used to support the analysis of a…