10 papers
SymStep: Symbolic Step Verification for Logical Reasoning
Aida Usmanova, Rui Gao, Dilshod Azizov +2
Chain-of-thought (CoT) prompting can fail severely on constraint-dense logical reasoning tasks, where unverified errors accumulate silently across steps. We introduce SymStep: an L…
A Multi-View Media Profiling Suite: Resources, Evaluation, and Analysis
Muhammad Arslan Manzoor, Dilshod Azizov, Daniil Orel +4
News outlets shape public opinion at a scale that makes automated detection of political bias and factuality essential. However, the field still lacks unified resources, comprehens…
AICD Bench: A Challenging Benchmark for AI-Generated Code Detection
Daniil Orel, Dilshod Azizov, Indraneil Paul +3
Large language models (LLMs) are increasingly capable of generating functional source code, raising concerns about authorship, accountability, and security. While detecting AI-gene…
Evolution of AI in Education: Agentic Workflows
Firuz Kamalov, David Santandreu Calonge, Linda Smail +4
The primary goal of this study is to analyze agentic workflows in education according to the proposed four major technological paradigms: reflection, planning, tool use, and multi-…
CoTaP: Compliant Task Pipeline and Reinforcement Learning of Its Controller with Compliance Modulation
Zewen He, Chenyuan Chen, Dilshod Azizov +1
Humanoid whole-body locomotion control is a critical approach for humanoid robots to leverage their inherent advantages. Learning-based control methods derived from retargeted huma…
CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings
Daniil Orel, Dilshod Azizov, Preslav Nakov
Large language models (LLMs) have revolutionized code generation, automating programming with remarkable efficiency. However, these advancements challenge programming skills, ethic…