3 citations · 7 across the 13 of their papers we have counts for
20 papers
Test-time Conditional Text-to-Image Synthesis Using Diffusion Models
Tripti Shukla, Srikrishna Karanam, Balaji Vasan Srinivasan
We consider the problem of conditional text-to-image synthesis with diffusion models. Most recent works need to either finetune specific parts of the base diffusion model or introd…
Training-free Color-Style Disentanglement for Constrained Text-to-Image Synthesis
Aishwarya Agarwal, Srikrishna Karanam, Balaji Vasan Srinivasan
We consider the problem of independently, in a disentangled fashion, controlling the outputs of text-to-image diffusion models with color and style attributes of a user-supplied re…
AlignIT: Enhancing Prompt Alignment in Customization of Text-to-Image Models
Aishwarya Agarwal, Srikrishna Karanam, Balaji Vasan Srinivasan
We consider the problem of customizing text-to-image diffusion models with user-supplied reference images. Given new prompts, the existing methods can capture the key concept from…
An Image is Worth Multiple Words: Multi-attribute Inversion for Constrained Text-to-Image Synthesis
Aishwarya Agarwal, Srikrishna Karanam, Tripti Shukla +1
We consider the problem of constraining diffusion model outputs with a user-supplied reference image. Our key objective is to extract multiple attributes (e.g., color, object, layo…
Iterative Multi-granular Image Editing using Diffusion Models
K J Joseph, Prateksha Udhayanan, Tripti Shukla +4
Recent advances in text-guided image synthesis has dramatically changed how creative professionals generate artistic and aesthetically pleasing visual assets. To fully support such…
Learning with Multi-modal Gradient Attention for Explainable Composed Image Retrieval
Prateksha Udhayanan, Srikrishna Karanam, Balaji Vasan Srinivasan
We consider the problem of composed image retrieval that takes an input query consisting of an image and a modification text indicating the desired changes to be made on the image…