1 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 1 cited
HAAV: Hierarchical Aggregation of Augmented Views for Image Captioning
Chia-Wen Kuo, Zsolt Kira
A great deal of progress has been made in image captioning, driven by research into how to encode the image using pre-trained models. This includes visual encodings (e.g. image gri…
cs.CV2023★ 1 cited
CLIP-GCD: Simple Language Guided Generalized Category Discovery
Rabah Ouldnoughi, Chia-Wen Kuo, Zsolt Kira
Generalized Category Discovery (GCD) requires a model to both classify known categories and cluster unknown categories in unlabeled data. Prior methods leveraged self-supervised pr…