activity
20222026
most citedRobust Reinforcement Learning on Graphs for Logistics optimization

2 citations · 5 across the 12 of their papers we have counts for

collaborators
Showing cs.AIShow all

10 papers · 1 filter

cs.AI2026

How Robust Are Automated Fact-Checking Systems? A Cross-Benchmark Evaluation

Aida Usmanova, Zangir Iklassov, Markus Leippold +1

Automated fact-checking (AFC) systems retrieve evidence and predict claim veracity, yet evaluations omit simple baselines, systems are developed for a single benchmark and cannot b…

cs.AI2026

SymStep: Symbolic Step Verification for Logical Reasoning

Aida Usmanova, Rui Gao, Dilshod Azizov +2

Chain-of-thought (CoT) prompting can fail severely on constraint-dense logical reasoning tasks, where unverified errors accumulate silently across steps. We introduce SymStep: an L…

cs.AI2026

Einstein World Models

Munachiso Samuel Nwadike, Zangir Iklassov, Ali Mekky +2

Does intelligence require the ability to reason about phenomena beyond direct experience? It is natural to suspect that some complex thought cannot be captured through language alo…

cs.AI2026

Measuring AI Reasoning: A Guide for Researchers

Munachiso Samuel Nwadike, Zangir Iklassov, Kareem Ali +2

In this paper, we offer a guide for researchers on evaluating reasoning in language models, building the case that reasoning should be assessed through evidence of adaptive, multi-…

cs.AI2025★ 1 cited

The AI Data Scientist

Farkhad Akimov, Munachiso Samuel Nwadike, Zangir Iklassov +1

Imagine decision-makers uploading data and, within minutes, receiving clear, actionable insights delivered straight to their fingertips. That is the promise of the AI Data Scientis…

cs.AI2025

SVRPBench: A Realistic Benchmark for Stochastic Vehicle Routing Problem

Ahmed Heakl, Yahia Salaheldin Shaaban, Martin Takac +2

Robust routing under uncertainty is central to real-world logistics, yet most benchmarks assume static, idealized settings. We present SVRPBench, the first open benchmark to captur…