most citedLearning a Recurrent Visual Representation for Image Caption Generation

179 citations · 291 across the 5 of their papers we have counts for

collaborators

5 papers

cs.LG202328 cited

Reducing SO(3) Convolutions to SO(2) for Efficient Equivariant GNNs

Saro Passaro, C. Lawrence Zitnick

Graph neural networks that model 3D data, such as point clouds or atoms, are typically desired to be equivariant, i.e., equivariant to 3D rotations. Unfortunately equivaria…

cs.CV2014179 cited

Learning a Recurrent Visual Representation for Image Caption Generation

Xinlei Chen, C. Lawrence Zitnick

In this paper we explore the bi-directional mapping between images and their sentence-based descriptions. We propose learning this mapping using a recurrent neural network. Unlike…

cs.CV20145 cited

Collecting Image Description Datasets using Crowdsourcing

Ramakrishna Vedantam, C. Lawrence Zitnick, Devi Parikh

We describe our two new datasets with images described by humans. Both the datasets were collected using Amazon Mechanical Turk, a crowdsourcing platform. The two datasets contain…

cs.CV201461 cited

CIDEr: Consensus-based Image Description Evaluation

Ramakrishna Vedantam, C. Lawrence Zitnick, Devi Parikh

Automatically describing an image with a sentence is a long-standing challenge in computer vision and natural language processing. Due to recent progress in object detection, attri…

cs.CV201418 cited

Fast Edge Detection Using Structured Forests

Piotr Dollár, C. Lawrence Zitnick

Edge detection is a critical component of many vision systems, including object detectors and image segmentation algorithms. Patches of edges exhibit well-known forms of local stru…