184 citations · 200 across the 7 of their papers we have counts for
8 papers · 1 filter
WikiMulti: a Corpus for Cross-Lingual Summarization
Pavel Tikhonov, Valentin Malykh
Cross-lingual summarization (CLS) is the task to produce a summary in one particular language for a source document in a different language. We introduce WikiMulti - a new dataset…
Russian SuperGLUE 1.1: Revising the Lessons not Learned by Russian NLP models
Alena Fenogenova, Maria Tikhonova, Vladislav Mikhailov +6
In the last year, new neural architectures and multilingual pre-trained models have been released for Russian, which led to performance evaluation problems across a range of langua…
A Single Example Can Improve Zero-Shot Data Generation
Pavel Burnyshev, Valentin Malykh, Andrey Bout +2
Sub-tasks of intent classification, such as robustness to distribution shift, adaptation to specific user groups and personalization, out-of-domain detection, require extensive and…
MOROCCO: Model Resource Comparison Framework
Valentin Malykh, Alexander Kukushkin, Ekaterina Artemova +3
The new generation of pre-trained NLP models push the SOTA to the new limits, but at the cost of computational resources, to the point that their use in real production environment…
Improving unsupervised neural aspect extraction for online discussions using out-of-domain classification
Anton Alekseev, Elena Tutubalina, Valentin Malykh +1
Deep learning architectures based on self-attention have recently achieved and surpassed state of the art results in the task of unsupervised aspect extraction and topic modeling.…
AspeRa: Aspect-based Rating Prediction Model
Sergey I. Nikolenko, Elena Tutubalina, Valentin Malykh +2
We propose a novel end-to-end Aspect-based Rating Prediction model (AspeRa) that estimates user rating based on review texts for the items and at the same time discovers coherent a…