activity
20192024
most citedImproving generation quality of pointer networks via guided attention

3 citations · 7 across the 13 of their papers we have counts for

collaborators

20 papers

cs.CV2024

Test-time Conditional Text-to-Image Synthesis Using Diffusion Models

Tripti Shukla, Srikrishna Karanam, Balaji Vasan Srinivasan

We consider the problem of conditional text-to-image synthesis with diffusion models. Most recent works need to either finetune specific parts of the base diffusion model or introd…

cs.CV2024

Training-free Color-Style Disentanglement for Constrained Text-to-Image Synthesis

Aishwarya Agarwal, Srikrishna Karanam, Balaji Vasan Srinivasan

We consider the problem of independently, in a disentangled fashion, controlling the outputs of text-to-image diffusion models with color and style attributes of a user-supplied re…

cs.CV2024

AlignIT: Enhancing Prompt Alignment in Customization of Text-to-Image Models

Aishwarya Agarwal, Srikrishna Karanam, Balaji Vasan Srinivasan

We consider the problem of customizing text-to-image diffusion models with user-supplied reference images. Given new prompts, the existing methods can capture the key concept from…

cs.CV20232 cited

An Image is Worth Multiple Words: Multi-attribute Inversion for Constrained Text-to-Image Synthesis

Aishwarya Agarwal, Srikrishna Karanam, Tripti Shukla +1

We consider the problem of constraining diffusion model outputs with a user-supplied reference image. Our key objective is to extract multiple attributes (e.g., color, object, layo…

cs.CV20231 cited

Iterative Multi-granular Image Editing using Diffusion Models

K J Joseph, Prateksha Udhayanan, Tripti Shukla +4

Recent advances in text-guided image synthesis has dramatically changed how creative professionals generate artistic and aesthetically pleasing visual assets. To fully support such…

cs.CV2023

Learning with Multi-modal Gradient Attention for Explainable Composed Image Retrieval

Prateksha Udhayanan, Srikrishna Karanam, Balaji Vasan Srinivasan

We consider the problem of composed image retrieval that takes an input query consisting of an image and a modification text indicating the desired changes to be made on the image…