2 papers
cs.AI2024
Multi-Modal Dialogue State Tracking for Playing GuessWhich Game
Wei Pang, Ruixue Duan, Jinfu Yang +1
GuessWhich is an engaging visual dialogue game that involves interaction between a Questioner Bot (QBot) and an Answer Bot (ABot) in the context of image-guessing. In this game, QB…
cs.AI2024
Enhancing Visual Dialog State Tracking through Iterative Object-Entity Alignment in Multi-Round Conversations
Wei Pang, Ruixue Duan, Jinfu Yang +1
Visual Dialog (VD) is a task where an agent answers a series of image-related questions based on a multi-round dialog history. However, previous VD methods often treat the entire d…