Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Semantic Document Derendering: SVG Reconstruction via Vision-Language Modeling
Adam Hazimeh, Ke Wang, Mark Collier +3
Multimedia documents such as slide presentations and posters are designed to be interactive and easy to modify. Yet, they are often distributed in a static raster format, which lim…
cs.CV2024
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
Gilles Baechler, Srinivas Sunkara, Maria Wang +7
Screen user interfaces (UIs) and infographics, sharing similar visual language and design principles, play important roles in human communication and human-machine interaction. We…