42 citations · 77 across the 6 of their papers we have counts for
Showing 2021 · cs.CVShow all
2 papers · 2 filters
cs.CV2021★ 11 cited
LatteGAN: Visually Guided Language Attention for Multi-Turn Text-Conditioned Image Manipulation
Shoya Matsumori, Yuki Abe, Kosuke Shingyouchi +2
Text-guided image manipulation tasks have recently gained attention in the vision-and-language community. While most of the prior studies focused on single-turn manipulation, our g…
cs.CV2021
Unified Questioner Transformer for Descriptive Question Generation in Goal-Oriented Visual Dialogue
Shoya Matsumori, Kosuke Shingyouchi, Yuki Abe +3
Building an interactive artificial intelligence that can ask questions about the real world is one of the biggest challenges for vision and language problems. In particular, goal-o…