596 citations · 652 across the 18 of their papers we have counts for
41 papers
Scene Text Recognition with Semantics
Joshua Cesare Placidi, Yishu Miao, Zixu Wang +1
Scene Text Recognition (STR) models have achieved high performance in recent years on benchmark datasets where text images are presented with minimal noise. Traditional STR recogni…
A Survey of Online Hate Speech through the Causal Lens
Antigoni-Maria Founta, Lucia Specia
The societal issue of digital hostility has previously attracted a lot of attention. The topic counts an ample body of literature, yet remains prominent and challenging as ever due…
BERTGEN: Multi-task Generation through BERT
Faidon Mitzalis, Ozan Caglayan, Pranava Madhyastha +1
We present BERTGEN, a novel generative, decoder-only model which extends BERT by fusing multimodal and multilingual pretrained models VL-BERT and M-BERT, respectively. BERTGEN is a…
Cross-Modal Generative Augmentation for Visual Question Answering
Zixu Wang, Yishu Miao, Lucia Specia
Data augmentation has been shown to effectively improve the performance of multimodal machine learning models. This paper introduces a generative model for data augmentation by lev…
Exploring Supervised and Unsupervised Rewards in Machine Translation
Julia Ive, Zixu Wang, Marina Fomicheva +1
Reinforcement Learning (RL) is a powerful framework to address the discrepancy between loss functions used during training and the final evaluation metrics to be used at test time.…
Exploiting Multimodal Reinforcement Learning for Simultaneous Machine Translation
Julia Ive, Andy Mingren Li, Yishu Miao +3
This paper addresses the problem of simultaneous machine translation (SiMT) by exploring two main concepts: (a) adaptive policies to learn a good trade-off between high translation…