82 citations · 87 across the 5 of their papers we have counts for
3 papers · 1 filter
Going for GOAL: A Resource for Grounded Football Commentaries
Alessandro Suglia, José Lopes, Emanuele Bastianelli +6
Recent video+language datasets cover domains where the interaction is highly structured, such as instructional videos, or where the interaction is scripted, such as TV shows. Both…
History for Visual Dialog: Do we really need it?
Shubham Agarwal, Trung Bui, Joon-Young Lee +2
Visual Dialog involves "understanding" the dialog history (what has been discussed previously) and the current question (what is asked), in addition to grounding information in the…
Ensemble based discriminative models for Visual Dialog Challenge 2018
Shubham Agarwal, Raghav Goyal
This manuscript describes our approach for the Visual Dialog Challenge 2018. We use an ensemble of three discriminative models with different encoders and decoders for our final su…