1 citations · 1 across the 1 of their papers we have counts for
3 papers
Subspace-Boosted Model Merging
Ronald Skorobogat, Karsten Roth, Mariana-Iuliana Georgescu
Model merging enables the combination of multiple specialized expert models into a single model capable of performing multiple tasks. However, the benefits of merging an increasing…
WikiBigEdit: Understanding the Limits of Lifelong Knowledge Editing in LLMs
Lukas Thede, Karsten Roth, Matthias Bethge +2
Keeping large language models factually up-to-date is crucial for deployment, yet costly retraining remains a challenge. Knowledge editing offers a promising alternative, but metho…
Context-Aware Multimodal Pretraining
Karsten Roth, Zeynep Akata, Dima Damen +2
Large-scale multimodal representation learning successfully optimizes for zero-shot transfer at test time. Yet the standard pretraining paradigm (contrastive learning on large amou…