179 citations · 291 across the 5 of their papers we have counts for
5 papers
Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNs
Saro Passaro, C. Lawrence Zitnick
Graph neural networks that model 3D data, such as point clouds or atoms, are typically desired to be equivariant, i.e., equivariant to 3D rotations. Unfortunately equivaria…
Learning a Recurrent Visual Representation for Image Caption Generation
Xinlei Chen, C. Lawrence Zitnick
In this paper we explore the bi-directional mapping between images and their sentence-based descriptions. We propose learning this mapping using a recurrent neural network. Unlike…
Collecting Image Description Datasets using Crowdsourcing
Ramakrishna Vedantam, C. Lawrence Zitnick, Devi Parikh
We describe our two new datasets with images described by humans. Both the datasets were collected using Amazon Mechanical Turk, a crowdsourcing platform. The two datasets contain…
CIDEr: Consensus-based Image Description Evaluation
Ramakrishna Vedantam, C. Lawrence Zitnick, Devi Parikh
Automatically describing an image with a sentence is a long-standing challenge in computer vision and natural language processing. Due to recent progress in object detection, attri…
Fast Edge Detection Using Structured Forests
Piotr Dollár, C. Lawrence Zitnick
Edge detection is a critical component of many vision systems, including object detectors and image segmentation algorithms. Patches of edges exhibit well-known forms of local stru…