Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
AsyncPatch Diffusion: spatially-flexible image generation
Samuele Papa, Valentin De Bortoli, Guillaume Couairon +3
Standard diffusion models corrupt an entire sample with a single shared noise level, forcing all spatial regions to follow the same denoising trajectory. We introduce AsyncPatch Di…
cs.CV2024
NARAIM: Native Aspect Ratio Autoregressive Image Models
Daniel Gallo Fernández, Robert van der Klis, Răzvan-Andrei Matişan +4
While vision transformers are able to solve a wide variety of computer vision tasks, no pre-training method has yet demonstrated the same scaling laws as observed in language model…