6 papers
Understanding Gender and Racial Disparities in Image Recognition Models
Rohan Mahadev, Anindya Chakravarti
Large scale image classification models trained on top of popular datasets such as Imagenet have shown to have a distributional skew which leads to disparities in prediction accura…
Improving Visual Recognition using Ambient Sound for Supervision
Rohan Mahadev, Hongyu Lu
Our brains combine vision and hearing to create a more elaborate interpretation of the world. When the visual input is insufficient, a rich panoply of sounds can be used to describ…
Demystifying Multi-Faceted Video Summarization: Tradeoff Between Diversity,Representation, Coverage and Importance
Vishal Kaushal, Rishabh Iyer, Khoshrav Doctor +6
This paper addresses automatic summarization of videos in a unified manner. In particular, we propose a framework for multi-faceted summarization for extractive, query base and ent…
Learning From Less Data: A Unified Data Subset Selection and Active Learning Framework for Computer Vision
Vishal Kaushal, Rishabh Iyer, Suraj Kothawade +3
Supervised machine learning based state-of-the-art computer vision techniques are in general data hungry. Their data curation poses the challenges of expensive human labeling, inad…
Vis-DSS: An Open-Source toolkit for Visual Data Selection and Summarization
Rishabh Iyer, Pratik Dubal, Kunal Dargan +3
With increasing amounts of visual data being created in the form of videos and images, visual data selection and summarization are becoming ever increasing problems. We present Vis…
Deployment of Customized Deep Learning based Video Analytics On Surveillance Cameras
Pratik Dubal, Rohan Mahadev, Suraj Kothawade +2
This paper demonstrates the effectiveness of our customized deep learning based video analytics system in various applications focused on security, safety, customer analytics and p…