162 citations · 310 across the 5 of their papers we have counts for
Showing 2016Show all
2 papers · 1 filter
cs.CL2016
Visual Storytelling
Ting-Hao, Huang, Francis Ferraro +13
We introduce the first dataset for sequential vision-to-language, and explore how this data may be used for the task of visual storytelling. The first release of this dataset, SIND…
cs.CL2016
Generating Natural Questions About an Image
Nasrin Mostafazadeh, Ishan Misra, Jacob Devlin +3
There has been an explosion of work in the vision & language community during the past few years from image captioning to video transcription, and answering questions about images.…