3 papers
cs.CL2024
Detecting Concrete Visual Tokens for Multimodal Machine Translation
Braeden Bowen, Vipin Vijayan, Scott Grigsby +2
The challenge of visual grounding and masking in multimodal machine translation (MMT) systems has encouraged varying approaches to the detection and selection of visually-grounded…
cs.CL2024
Adding Multimodal Capabilities to a Text-only Translation Model
Vipin Vijayan, Braeden Bowen, Scott Grigsby +2
While most current work in multimodal machine translation (MMT) uses the Multi30k dataset for training and evaluation, we find that the resulting models overfit to the Multi30k dat…
cs.CL2024
The Case for Evaluating Multimodal Translation Models on Text Datasets
Vipin Vijayan, Braeden Bowen, Scott Grigsby +2
A good evaluation framework should evaluate multimodal machine translation (MMT) models by measuring 1) their use of visual information to aid in the translation task and 2) their…