1 citations · 1 across the 1 of their papers we have counts for
1 paper
Jason Tang, Garrin McGoldrick, Marie Al-Ghossein +1
This paper explores the usage of multimodal image-to-text models to enhance text-based item retrieval. We propose utilizing pre-trained image captioning and tagging models, such as…