4 papers
Multi-Agentic System Leveraging Open-Source LLMs to Mitigate Disinformation Threats
Sebastian Kula, Martin Tamajka
In contemporary societies, the threat of disinformation has reached alarming levels, exacerbated by the proliferation of electronic communication, social media, and advancements in…
MultiCW: A Large-Scale Balanced Benchmark Dataset for Training Robust Check-Worthiness Detection Models
Martin Hyben, Sebastian Kula, Jan Cegin +3
Large Language Models (LLMs) are beginning to reshape how media professionals verify information, yet automated support for detecting check-worthy claims a key step in the fact-che…
DelTriC: A Novel Clustering Method with Accurate Outlier
Tomas Javurek, Michal Gregor, Sebastian Kula +1
The paper introduces DelTriC (Delaunay Triangulation Clustering), a clustering algorithm which integrates PCA/UMAP-based projection, Delaunay triangulation, and a novel back-projec…
Multilingual and Multi-topical Benchmark of Fine-tuned Language models and Large Language Models for Check-Worthy Claim Detection
Martin Hyben, Sebastian Kula, Ivan Srba +2
This study compares the performance of (1) fine-tuned language models and (2) large language models on the task of check-worthy claim detection. For the purpose of the comparison w…