Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
How Robust Are Automated Fact-Checking Systems? A Cross-Benchmark Evaluation
Aida Usmanova, Zangir Iklassov, Markus Leippold +1
Automated fact-checking (AFC) systems retrieve evidence and predict claim veracity, yet evaluations omit simple baselines, systems are developed for a single benchmark and cannot b…
cs.AI2026
SymStep: Symbolic Step Verification for Logical Reasoning
Aida Usmanova, Rui Gao, Dilshod Azizov +2
Chain-of-thought (CoT) prompting can fail severely on constraint-dense logical reasoning tasks, where unverified errors accumulate silently across steps. We introduce SymStep: an L…
cs.AI2025
Automating SPARQL Query Translations between DBpedia and Wikidata
Malte Christian Bartels, Debayan Banerjee, Ricardo Usbeck
This paper investigates whether state-of-the-art Large Language Models (LLMs) can automatically translate SPARQL between popular Knowledge Graph (KG) schemas. We focus on translati…