4 papers
Iterative Layer-wise Distillation for Efficient Compression of Large Language Models
Grigory Kovalev, Mikhail Tikhomirov
This work investigates distillation methods for large language models (LLMs) with the goal of developing compact models that preserve high performance. Several existing approaches…
Wikipedia-based Datasets in Russian Information Retrieval Benchmark RusBEIR
Grigory Kovalev, Natalia Loukachevitch, Mikhail Tikhomirov +2
In this paper, we present a novel series of Russian information retrieval datasets constructed from the "Did you know..." section of Russian Wikipedia. Our datasets support a range…
Building Russian Benchmark for Evaluation of Information Retrieval Models
Grigory Kovalev, Mikhail Tikhomirov, Evgeny Kozhevnikov +2
We introduce RusBEIR, a comprehensive benchmark designed for zero-shot evaluation of information retrieval (IR) models in the Russian language. Comprising 17 datasets from various…
RuOpinionNE-2024: Extraction of Opinion Tuples from Russian News Texts
Natalia Loukachevitch, Natalia Tkachenko, Anna Lapanitsyna +2
In this paper, we introduce the Dialogue Evaluation shared task on extraction of structured opinions from Russian news texts. The task of the contest is to extract opinion tuples f…