230 citations · 232 across the 2 of their papers we have counts for
4 papers
Towards Fairer Datasets: Filtering and Balancing the Distribution of the People Subtree in the ImageNet Hierarchy
Kaiyu Yang, Klint Qinami, Li Fei-Fei +2
Computer vision technology is being used by many but remains representative of only a few. People have reported misbehavior of computer vision models, including offensive predictio…
Compositional Temporal Visual Grounding of Natural Language Event Descriptions
Jonathan C. Stroud, Ryan McCaffrey, Rada Mihalcea +2
Temporal grounding entails establishing a correspondence between natural language event descriptions and their visual depictions. Compositional modeling becomes central: we first g…
SpatialSense: An Adversarially Crowdsourced Benchmark for Spatial Relation Recognition
Kaiyu Yang, Olga Russakovsky, Jia Deng
Understanding the spatial relations between objects in images is a surprisingly challenging task. A chair may be "behind" a person even if it appears to the left of the person in t…
CornerNet-Lite: Efficient Keypoint Based Object Detection
Hei Law, Yun Teng, Olga Russakovsky +1
Keypoint-based methods are a relatively new paradigm in object detection, eliminating the need for anchor boxes and offering a simplified detection framework. Keypoint-based Corner…