1 paper · 1 filter
Stefan Cassar, Adrian Muscat, Dylan Seychell
Vision and language tasks such as Visual Relation Detection and Visual Question Answering benefit from semantic features that afford proper grounding of language. The 3D depth of o…