activity
20192024
most citedImproving Conditioning in Context-Aware Sequence to Sequence Models

9 citations · 21 across the 10 of their papers we have counts for

collaborators
Showing cs.CLShow all

7 papers · 1 filter

cs.CL20233 cited

mmT5: Modular Multilingual Pre-Training Solves Source Language Hallucinations

Jonas Pfeiffer, Francesco Piccinno, Massimo Nicosia +3

Multilingual sequence-to-sequence models perform poorly with increased language coverage and fail to consistently generate text in the correct target language in few-shot settings.…

cs.CL20232 cited

Serial Contrastive Knowledge Distillation for Continual Few-shot Relation Extraction

Xinyi Wang, Zitao Wang, Wei Hu

Continual few-shot relation extraction (RE) aims to continuously train a model for new relations with few labeled training data, of which the major challenges are the catastrophic…

cs.CL20222 cited

Enhancing Document-level Relation Extraction by Entity Knowledge Injection

Xinyi Wang, Zitao Wang, Weijian Sun +1

Document-level relation extraction (RE) aims to identify the relations between entities throughout an entire document. It needs complex reasoning skills to synthesize various knowl…

cs.CL2021

Efficient Test Time Adapter Ensembling for Low-resource Language Varieties

Xinyi Wang, Yulia Tsvetkov, Sebastian Ruder +1

Adapters are light-weight modules that allow parameter-efficient fine-tuning of pretrained models. Specialized language and task adapters have recently been proposed to facilitate…

cs.CL20212 cited

Multi-view Subword Regularization

Xinyi Wang, Sebastian Ruder, Graham Neubig

Multilingual pretrained representations generally rely on subword segmentation algorithms to create a shared multilingual vocabulary. However, standard heuristic algorithms often l…

cs.CL2020

Improving Target-side Lexical Transfer in Multilingual Neural Machine Translation

Luyu Gao, Xinyi Wang, Graham Neubig

To improve the performance of Neural Machine Translation~(NMT) for low-resource languages~(LRL), one effective strategy is to leverage parallel data from a related high-resource la…