1 paper
Ravi Shekhar, Ece Takmaz, Raquel Fernández +1
The multimodal models used in the emerging field at the intersection of computational linguistics and computer vision implement the bottom-up processing of the `Hub and Spoke' arch…