2 papers
cs.AI2025
OneFlow: Concurrent Mixed-Modal and Interleaved Generation with Edit Flows
John Nguyen, Marton Havasi, Tariq Berrada +2
We present OneFlow, the first non-autoregressive multimodal model that enables variable-length and concurrent mixed-modal generation. Unlike autoregressive models that enforce rigi…
cs.CV2025
Boosting Latent Diffusion with Perceptual Objectives
Tariq Berrada, Pietro Astolfi, Melissa Hall +6
Latent diffusion models (LDMs) power state-of-the-art high-resolution generative image models. LDMs learn the data distribution in the latent space of an autoencoder (AE) and produ…