2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CL2022★ 2 cited
Do BERTs Learn to Use Browser User Interface? Exploring Multi-Step Tasks with Unified Vision-and-Language BERTs
Taichi Iki, Akiko Aizawa
Pre-trained Transformers are good foundations for unified multi-task models owing to their task-agnostic representation. Pre-trained Transformers are often combined with text-to-te…
cs.CL2021
Effect of Visual Extensions on Natural Language Understanding in Vision-and-Language Models
Taichi Iki, Akiko Aizawa
A method for creating a vision-and-language (V&L) model is to extend a language model through structural modifications and V&L pre-training. Such an extension aims to make a V&L mo…