activity
20162023
most citedAutomatic Spanish Translation of the SQuAD Dataset for Multilingual Question Answering

42 citations · 80 across the 12 of their papers we have counts for

collaborators

20 papers

cs.CL2023

Speech Translation with Foundation Models and Optimal Transport: UPC at IWSLT23

Ioannis Tsiamas, Gerard I. Gállego, José A. R. Fonollosa +1

This paper describes the submission of the UPC Machine Translation group to the IWSLT 2023 Offline Speech Translation task. Our Speech Translation systems utilize foundation models…

cs.CL2022★ 1 cited

SegAugment: Maximizing the Utility of Speech Translation Data with Segmentation-based Augmentations

Ioannis Tsiamas, José A. R. Fonollosa, Marta R. Costa-jussà

End-to-end Speech Translation is hindered by a lack of available data resources. While most of them are based on documents, a sentence-level version is available, which is however…

cs.CL2022★ 1 cited

Efficient Speech Translation with Dynamic Latent Perceivers

Ioannis Tsiamas, Gerard I. Gállego, José A. R. Fonollosa +1

Transformers have been the dominant architecture for Speech Translation in recent years, achieving significant improvements in translation quality. Since speech signals are longer…

cs.SD2022★ 3 cited

SHAS: Approaching optimal Segmentation for End-to-End Speech Translation

Ioannis Tsiamas, Gerard I. Gállego, José A. R. Fonollosa +1

Speech translation models are unable to directly process long audios, like TED talks, which have to be split into shorter segments. Speech translation datasets provide manual segme…

cs.CL2021

End-to-End Speech Translation with Pre-trained Models and Adapters: UPC at IWSLT 2021

Gerard I. Gállego, Ioannis Tsiamas, Carlos Escolano +2

This paper describes the submission to the IWSLT 2021 offline speech translation task by the UPC Machine Translation group. The task consists of building a system capable of transl…

cs.CL2021★ 2 cited

Sparsely Factored Neural Machine Translation

Noe Casas, Jose A. R. Fonollosa, Marta R. Costa-jussà

The standard approach to incorporate linguistic information to neural machine translation systems consists in maintaining separate vocabularies for each of the annotated features t…