2 papers
cs.CV2025
InfSplign: Inference-Time Spatial Alignment of Text-to-Image Diffusion Models
Sarah Rastegar, Violeta Chatalbasheva, Sieger Falkena +5
Text-to-image (T2I) diffusion models generate high-quality images but often fail to capture the spatial relations specified in text prompts. This limitation can be traced to two fa…
cs.CV2025
Data-Efficient Challenges in Visual Inductive Priors: A Retrospective
Robert-Jan Bruintjes, Attila Lengyel, Osman Semih Kayhan +4
Deep Learning requires large amounts of data to train models that work well. In data-deficient settings, performance can be degraded. We investigate which Deep Learning methods ben…