1 citations · 1 across the 1 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
GraPE: A Generate-Plan-Edit Framework for Compositional T2I Synthesis
Ashish Goswami, Satyam Kumar Modi, Santhosh Rishi Deshineni +3
Text-to-image (T2I) generation has seen significant progress with diffusion models, enabling generation of photo-realistic images from text prompts. Despite this progress, existing…
cs.CV2024
Learning Disentangled Representation in Object-Centric Models for Visual Dynamics Prediction via Transformers
Sanket Gandhi, Atul, Samanyu Mahajan +4
Recent work has shown that object-centric representations can greatly help improve the accuracy of learning dynamics while also bringing interpretability. In this work, we take thi…