1 citations · 2 across the 9 of their papers we have counts for
3 papers · 2 filters
FODVid: Flow-guided Object Discovery in Videos
Silky Singh, Shripad Deshmukh, Mausoom Sarkar +3
Segmentation of objects in a video is challenging due to the nuances such as motion blurring, parallax, occlusions, changes in illumination, etc. Instead of addressing these nuance…
Parameter Efficient Local Implicit Image Function Network for Face Segmentation
Mausoom Sarkar, Nikitha SR, Mayur Hemani +2
Face parsing is defined as the per-pixel labeling of images containing human faces. The labels are defined to identify key facial regions like eyes, lips, nose, hair, etc. In this…
DeAR: Debiasing Vision-Language Models with Additive Residuals
Ashish Seth, Mayur Hemani, Chirag Agarwal
Large pre-trained vision-language models (VLMs) reduce the time for developing predictive models for various vision-grounded language downstream tasks by providing rich, adaptable…