5 papers
LLM Agents Factory: Retrieval of Domain-Specific LLM Agents
Vitalii Belov, Artyom Sosedka, Andrey Sakhovskiy +3
Large language model (LLM) agents improve task performance by decomposing problems into role-specialized behaviors. However, their practical deployment is often limited by the comp…
The Chronicles of RiDiC: Generating Datasets with Controlled Popularity Distribution for Long-form Factuality Evaluation
Pavel Braslavski, Dmitrii Iarosh, Nikita Sushko +4
We present a configurable pipeline for generating multilingual sets of entities with specified characteristics, such as domain, geographical location and popularity, using data fro…
RuCCoD: Towards Automated ICD Coding in Russian
Aleksandr Nesterov, Andrey Sakhovskiy, Ivan Sviridov +5
This study investigates the feasibility of automating clinical coding in Russian, a language with limited biomedical resources. We present a new dataset for ICD coding, which inclu…
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
Andrey Sakhovskiy, Elena Tutubalina
In recent years, there has been substantial progress in using pretrained Language Models (LMs) on a range of tasks aimed at improving the understanding of biomedical texts. Nonethe…
Overview of BioASQ 2025: The Thirteenth BioASQ Challenge on Large-Scale Biomedical Semantic Indexing and Question Answering
Anastasios Nentidis, Georgios Katsimpras, Anastasia Krithara +17
This is an overview of the thirteenth edition of the BioASQ challenge in the context of the Conference and Labs of the Evaluation Forum (CLEF) 2025. BioASQ is a series of internati…