activity
20242026
collaborators

7 papers

cs.CL2026

Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resource Languages

Saeed Almheiri, Bilal Elbouardi, Salsabila Zahirah Pranida +16

Idiomatic expressions pose a major challenge for multilingual NLP because their meanings shift between figurative and literal usage, often requiring context for accurate interpreta…

cs.CL2026

Leveraging LLM Parametric Knowledge for Fact Checking without Retrieval

Artem Vazhentsev, Maria Marina, Daniil Moskovskiy +8

Trustworthiness is a core research challenge for agentic AI systems built on Large Language Models (LLMs). To enhance trust, natural language claims from diverse sources, including…

cs.CL2025

Do I look like a `cat.n.01` to you? A Taxonomy Image Generation Benchmark

Viktor Moskvoretskii, Alina Lobanova, Ekaterina Neminova +3

This paper explores the feasibility of using text-to-image models in a zero-shot setup to generate images for taxonomy concepts. While text-based methods for taxonomy enrichment ar…

cs.CL2025

Self-Taught Self-Correction for Small Language Models

Viktor Moskvoretskii, Chris Biemann, Irina Nikishina

Although large language models (LLMs) have achieved remarkable performance across various tasks, they remain prone to errors. A key challenge is enabling them to self-correct. Whil…

cs.CL2025

Argument-Based Comparative Question Answering Evaluation Benchmark

Irina Nikishina, Saba Anwar, Nikolay Dolgov +6

In this paper, we aim to solve the problems standing in the way of automatic comparative question answering. To this end, we propose an evaluation framework to assess the quality o…

cs.IR2025

Creating a Taxonomy for Retrieval Augmented Generation Applications

Irina Nikishina, Özge Sevgili, Mahei Manhai Li +2

In this research, we develop a taxonomy to conceptualize a comprehensive overview of the constituting characteristics that define retrieval augmented generation (RAG) applications,…