activity
20182022
most citedGRIT: General Robust Image Task Benchmark

11 citations · 11 across the 1 of their papers we have counts for

collaborators

7 papers

cs.CV202211 cited

GRIT: General Robust Image Task Benchmark

Tanmay Gupta, Ryan Marten, Aniruddha Kembhavi +1

Computer vision models excel at making predictions when the test distribution closely resembles the training distribution. Such models have yet to match the ability of biological v…

cs.CV2021

Visual Semantic Role Labeling for Video Understanding

Arka Sadhu, Tanmay Gupta, Mark Yatskar +2

We propose a new framework for understanding and representing related salient events in a video using visual semantic role labeling. We represent videos as a set of related events,…

cs.LG2020

Learning Curves for Analysis of Deep Networks

Derek Hoiem, Tanmay Gupta, Zhizhong Li +1

Learning curves model a classifier's test error as a function of the number of training samples. Prior works show that learning curves can be used to select model parameters and ex…

cs.CV2020

Contrastive Learning for Weakly Supervised Phrase Grounding

Tanmay Gupta, Arash Vahdat, Gal Chechik +3

Phrase grounding, the problem of associating image regions to caption words, is a crucial component of vision-language tasks. We show that phrase grounding can be learned by optimi…

cs.CV2019

ViCo: Word Embeddings from Visual Co-occurrences

Tanmay Gupta, Alexander Schwing, Derek Hoiem

We propose to learn word embeddings from visual co-occurrences. Two words co-occur visually if both words apply to the same image or image region. Specifically, we extract four typ…

cs.CV2018

No-Frills Human-Object Interaction Detection: Factorization, Layout Encodings, and Training Techniques

Tanmay Gupta, Alexander Schwing, Derek Hoiem

We show that for human-object interaction detection a relatively simple factorized model with appearance and layout encodings constructed from pre-trained object detectors outperfo…