Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Attentive multilayer fusion for vision transformers
Laure Ciernik, Marco Morik, Lukas Thede +4
With the rise of large-scale foundation models, efficiently adapting them to downstream tasks remains a central challenge. Linear probing, which freezes the backbone and trains a l…
cs.CV2024
ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise Optimization
Luca Eyring, Shyamgopal Karthik, Karsten Roth +2
Text-to-Image (T2I) models have made significant advancements in recent years, but they still struggle to accurately capture intricate details specified in complex compositional pr…