3 papers
cs.CV2026
PLACID: Identity-Preserving Multi-Object Compositing via Video Diffusion with Synthetic Trajectories
Gemma Canet Tarrés, Manel Baradad, Francesc Moreno-Noguer +1
Recent advances in generative AI have dramatically improved photorealistic image synthesis, yet they fall short for studio-level multi-object compositing. This task demands simulta…
cs.CV2025
Cost Savings from Automatic Quality Assessment of Generated Images
Xavier Giro-i-Nieto, Nefeli Andreou, Anqi Liang +3
Deep generative models have shown impressive progress in recent years, making it possible to produce high quality images with a simple text prompt or a reference image. However, st…
cs.CV2025
Separating Knowledge and Perception with Procedural Data
Adrián Rodríguez-Muñoz, Manel Baradad, Phillip Isola +1
We train representation models with procedural data only, and apply them on visual similarity, classification, and semantic segmentation tasks without further training by using vis…