Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Semantic Document Derendering: SVG Reconstruction via Vision-Language Modeling
Adam Hazimeh, Ke Wang, Mark Collier +3
Multimedia documents such as slide presentations and posters are designed to be interactive and easy to modify. Yet, they are often distributed in a static raster format, which lim…
cs.CV2025
Single-Input Multi-Output Model Merging: Leveraging Foundation Models for Dense Multi-Task Learning
Juan Garcia Giraldo, Nikolaos Dimitriadis, Ke Wang +1
Model merging is a flexible and computationally tractable approach to merge single-task checkpoints into a multi-task model. Prior work has solely focused on constrained multi-task…