activity
20122025
most citedCaffe: Convolutional Architecture for Fast Feature Embedding

4.3k citations · 12.1k across the 42 of their papers we have counts for

collaborators
Showing 2023Show all

12 papers · 1 filter

cs.CV20231 cited

Unsupervised Universal Image Segmentation

Dantong Niu, Xudong Wang, Xinyang Han +3

Several unsupervised image segmentation approaches have been proposed which eliminate the need for dense manually-annotated segmentation masks; current models separately handle eit…

cs.LG20238 cited

Tensor Trust: Interpretable Prompt Injection Attacks from an Online Game

Sam Toyer, Olivia Watkins, Ethan Adrian Mendes +9

While Large Language Models (LLMs) are increasingly being used in real-world applications, they remain vulnerable to prompt injection attacks: malicious third party prompts that su…

cs.CV202314 cited

Large Language Models are Visual Reasoning Coordinators

Liangyu Chen, Bo Li, Sheng Shen +5

Visual reasoning requires multimodal perception and commonsense cognition of the world. Recently, multiple vision-language models (VLMs) have been proposed with excellent commonsen…

cs.CV2023

CLAIR: Evaluating Image Captions with Large Language Models

David Chan, Suzanne Petryk, Joseph E. Gonzalez +2

The evaluation of machine-generated image captions poses an interesting yet persistent challenge. Effective evaluation measures must consider numerous dimensions of similarity, inc…

cs.CV202312 cited

Aligning Large Multimodal Models with Factually Augmented RLHF

Zhiqing Sun, Sheng Shen, Shengcao Cao +9

Large Multimodal Models (LMM) are built across modalities and the misalignment between two modalities can result in "hallucination", generating textual outputs that are not grounde…

cs.CV2023

VideoCutLER: Surprisingly Simple Unsupervised Video Instance Segmentation

Xudong Wang, Ishan Misra, Ziyun Zeng +2

Existing approaches to unsupervised video instance segmentation typically rely on motion estimates and experience difficulties tracking small or divergent motions. We present Video…