1 citations · 2 across the 5 of their papers we have counts for
8 papers · 1 filter
DesignSense: A Human Preference Dataset and Reward Modeling Framework for Graphic Layout Generation
Varun Gopal, Rishabh Jain, Aradhya Mathur +6
Graphic layouts serve as an important and engaging medium for visual communication across different channels. While recent layout generation models have demonstrated impressive cap…
EOPose : Exemplar-based object reposing using Generalized Pose Correspondences
Sarthak Mehrotra, Rishabh Jain, Mayur Hemani +2
Reposing objects in images has a myriad of applications, especially for e-commerce where several variants of product images need to be produced quickly. In this work, we leverage t…
FODVid: Flow-guided Object Discovery in Videos
Silky Singh, Shripad Deshmukh, Mausoom Sarkar +3
Segmentation of objects in a video is challenging due to the nuances such as motion blurring, parallax, occlusions, changes in illumination, etc. Instead of addressing these nuance…
Parameter Efficient Local Implicit Image Function Network for Face Segmentation
Mausoom Sarkar, Nikitha SR, Mayur Hemani +2
Face parsing is defined as the per-pixel labeling of images containing human faces. The labels are defined to identify key facial regions like eyes, lips, nose, hair, etc. In this…
DeAR: Debiasing Vision-Language Models with Additive Residuals
Ashish Seth, Mayur Hemani, Chirag Agarwal
Large pre-trained vision-language models (VLMs) reduce the time for developing predictive models for various vision-grounded language downstream tasks by providing rich, adaptable…
ZFlow: Gated Appearance Flow-based Virtual Try-on with 3D Priors
Ayush Chopra, Rishabh Jain, Mayur Hemani +1
Image-based virtual try-on involves synthesizing perceptually convincing images of a model wearing a particular garment and has garnered significant research interest due to its im…