1 paper
Tong Ye, Shijing Si, Jianzong Wang +3
The visual dialog task attempts to train an agent to answer multi-turn questions given an image, which requires the deep understanding of interactions between the image and dialog…