9 citations · 14 across the 6 of their papers we have counts for
4 papers · 1 filter
ALOHa: A New Measure for Hallucination in Captioning Models
Suzanne Petryk, David M. Chan, Anish Kachinthaya +4
Despite recent advances in multimodal pre-training for visual description, state-of-the-art models still produce captions containing errors, such as hallucinating objects not prese…
CLAIR: Evaluating Image Captions with Large Language Models
David Chan, Suzanne Petryk, Joseph E. Gonzalez +2
The evaluation of machine-generated image captions poses an interesting yet persistent challenge. Effective evaluation measures must consider numerous dimensions of similarity, inc…
LAVA: Language Audio Vision Alignment for Contrastive Video Pre-Training
Sumanth Gurram, Andy Fang, David Chan +1
Generating representations of video data is of key importance in advancing the field of machine perception. Most current techniques rely on hand-annotated data, which can be diffic…
An Embedding-Dynamic Approach to Self-supervised Learning
Suhong Moon, Domas Buracas, Seunghyun Park +2
A number of recent self-supervised learning methods have shown impressive performance on image classification and other tasks. A somewhat bewildering variety of techniques have bee…