1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.CL2022★ 1 cited
Finding Structural Knowledge in Multimodal-BERT
Victor Milewski, Miryam de Lhoneux, Marie-Francine Moens
In this work, we investigate the knowledge learned in the embeddings of multimodal-BERT models. More specifically, we probe their capabilities of storing the grammatical structure…
cs.CV2020
Are scene graphs good enough to improve Image Captioning?
Victor Milewski, Marie-Francine Moens, Iacer Calixto
Many top-performing image captioning models rely solely on object features computed with an object detection model to generate image descriptions. However, recent studies propose t…