7 citations · 11 across the 2 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
MuSViT: A Foundation Vision Model for Sheet Music Representation
Carlos Penarrubia, Antonio Rios-Vila, Eliseo Fuentes-Martinez +4
Foundation models have transformed vision and language processing by providing rich, reusable representations that transfer across diverse tasks. Sheet music, as a visual encoding…
cs.CV2024★ 7 cited
Self-Supervised Learning for Text Recognition: A Critical Survey
Carlos Penarrubia, Jose J. Valero-Mas, Jorge Calvo-Zaragoza
Text Recognition (TR) refers to the research area that focuses on retrieving textual information from images, a topic that has seen significant advancements in the last decade due…