1 paper · 1 filter
Hongchen Wei, Zhenzhong Chen
Pretrained visual-language models have demonstrated impressive zero-shot abilities in image captioning, when accompanied by hand-crafted prompts. Meanwhile, hand-crafted prompts ut…