3 papers
cs.CV2026
Injecting Image Guidance into Text-Conditioned Diffusion Models at Inference
Agata Żywot, Iason Skylitsis, Thijmen Nijdam +4
Text-to-image diffusion models like Stable Diffusion generate high-quality images from text, but lack a way to inject visual guidance (e.g. sketches, styles) at inference without r…
cs.CV2026
LEMON: a foundation model for nuclear morphology in Computational Pathology
Loïc Chadoutaud, Alice Blondel, Hana Feki +3
Computational pathology relies on effective representation learning to support cancer research and precision medicine. Although self-supervised learning has driven major progress a…
cs.SD2025
Linear RNNs for autoregressive generation of long music samples
Konrad Szewczyk, Daniel Gallo Fernández, James Townsend
Directly learning to generate audio waveforms in an autoregressive manner is a challenging task, due to the length of the raw sequences and the existence of important structure on…