25 citations · 43 across the 4 of their papers we have counts for
12 papers · 1 filter
Vārta: A Large-Scale Headline-Generation Dataset for Indic Languages
Rahul Aralikatte, Ziling Cheng, Sumanth Doddapaneni +1
We present Vārta, a large-scale multilingual dataset for headline generation in Indic languages. This dataset includes 41.8 million news articles in 14 different Indic languages (a…
Minimax and Neyman-Pearson Meta-Learning for Outlier Languages
Edoardo Maria Ponti, Rahul Aralikatte, Disha Shrivastava +2
Model-agnostic meta-learning (MAML) has been recently put forth as a strategy to learn resource-poor languages in a sample-efficient fashion. Nevertheless, the properties of these…
Itihasa: A large-scale corpus for Sanskrit to English translation
Rahul Aralikatte, Miryam de Lhoneux, Anoop Kunchukuttan +1
This work introduces Itihasa, a large-scale translation dataset containing 93,000 pairs of Sanskrit shlokas and their English translations. The shlokas are extracted from two India…
Focus Attention: Promoting Faithfulness and Diversity in Summarization
Rahul Aralikatte, Shashi Narayan, Joshua Maynez +2
Professional summaries are written with document-level information, such as the theme of the document, in mind. This is in contrast with most seq2seq decoders which simultaneously…
Joint Semantic Analysis with Document-Level Cross-Task Coherence Rewards
Rahul Aralikatte, Mostafa Abdou, Heather Lent +2
Coreference resolution and semantic role labeling are NLP tasks that capture different aspects of semantics, indicating respectively, which expressions refer to the same entity, an…
Rewarding Coreference Resolvers for Being Consistent with World Knowledge
Rahul Aralikatte, Heather Lent, Ana Valeria Gonzalez +5
Unresolved coreference is a bottleneck for relation extraction, and high-quality coreference resolvers may produce an output that makes it a lot easier to extract knowledge triples…