4 citations · 5 across the 3 of their papers we have counts for
1 paper · 1 filter
Rita Ramos, Desmond Elliott, Bruno Martins
Inspired by retrieval-augmented language generation and pretrained Vision and Language (V&L) encoders, we present a new approach to image captioning that generates sentences given…