6 papers
Adversarial Creation and Detection of AI-Generated Social Bot Content
Mykola Trokhymovych, Ricardo Baeza-Yates, Alessandro Flammini +2
The convergence of large language models and social bots allows malicious actors to manipulate the information ecosystem by generating human-like content at scale. Existing models…
An End-to-End Ukrainian RAG for Local Deployment. Optimized Hybrid Search and Lightweight Generation
Mykola Trokhymovych, Yana Oliinyk, Nazarii Nyzhnyk
This paper presents a highly efficient Retrieval-Augmented Generation (RAG) system built specifically for Ukrainian document question answering, which achieved 2nd place in the UNL…
Multilingual Reference Need Assessment System for Wikipedia
Aitolkyn Baigutanova, Francisco Navas, Pablo Aragon +5
Wikipedia is a critical source of information for millions of users across the Web. It serves as a key resource for large language models, search engines, question-answering system…
Hidden Persuasion: Detecting Manipulative Narratives on Social Media During the 2022 Russian Invasion of Ukraine
Kateryna Akhynko, Oleksandr Kosovan, Mykola Trokhymovych
This paper presents one of the top-performing solutions to the UNLP 2025 Shared Task on Detecting Manipulation in Social Media. The task focuses on detecting and classifying rhetor…
Graph-Linguistic Fusion: Using Language Models for Wikidata Vandalism Detection
Mykola Trokhymovych, Lydia Pintscher, Ricardo Baeza-Yates +1
We introduce a next-generation vandalism detection system for Wikidata, one of the largest open-source structured knowledge bases on the Web. Wikidata is highly complex: its items…
Characterizing Knowledge Manipulation in a Russian Wikipedia Fork
Mykola Trokhymovych, Oleksandr Kosovan, Nathan Forrester +3
Wikipedia is powered by MediaWiki, a free and open-source software that is also the infrastructure for many other wiki-based online encyclopedias. These include the recently launch…