4 papers
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
Grigory Kovalev, Natalia Loukachevitch, Mikhail Tikhomirov +2
In this paper, we present a novel series of Russian information retrieval datasets constructed from the "Did you know..." section of Russian Wikipedia. Our datasets support a range…
Methods for Recognizing Nested Terms
Igor Rozhkov, Natalia Loukachevitch
In this paper, we describe our participation in the RuTermEval competition devoted to extracting nested terms. We apply the Binder model, which was previously successfully applied…
Building Russian Benchmark for Evaluation of Information Retrieval Models
Grigory Kovalev, Mikhail Tikhomirov, Evgeny Kozhevnikov +2
We introduce RusBEIR, a comprehensive benchmark designed for zero-shot evaluation of information retrieval (IR) models in the Russian language. Comprising 17 datasets from various…
RuOpinionNE-2024: Extraction of Opinion Tuples from Russian News Texts
Natalia Loukachevitch, Natalia Tkachenko, Anna Lapanitsyna +2
In this paper, we introduce the Dialogue Evaluation shared task on extraction of structured opinions from Russian news texts. The task of the contest is to extract opinion tuples f…