13 citations · 17 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 13 cited
MultiCapCLIP: Auto-Encoding Prompts for Zero-Shot Multilingual Visual Captioning
Bang Yang, Fenglin Liu, Xian Wu +3
Supervised visual captioning models typically require a large scale of images or videos paired with descriptions in a specific language (i.e., the vision-caption pairs) for trainin…
cs.CV2023★ 4 cited
Customizing General-Purpose Foundation Models for Medical Report Generation
Bang Yang, Asif Raza, Yuexian Zou +1
Medical caption prediction which can be regarded as a task of medical report generation (MRG), requires the automatic generation of coherent and accurate captions for the given med…