3 papers
cs.SE2026
Autoresearch with Coding Agents: Generalizers and Metric-Maximizers on Quran Recitation Data
Nursultan Askarbekuly, Mohamad Al Mdfaa, Ahmed Helaly +2
Coding agents can now be left alone to improve software against a score. In this pattern--recently popularized as "autoresearch"--the agent receives a dataset, an evaluation script…
cs.AI2025
An Outcome-Based Educational Recommender System
Nursultan Askarbekuly, Timur Fayzrakhmanov, Sladjan BabarogiÄ +1
Most educational recommender systems are tuned and judged on click- or rating-based relevance, leaving their true pedagogical impact unclear. We introduce OBER-an Outcome-Based Edu…
cs.CL2025
Sacred or Synthetic? Evaluating LLM Reliability and Abstention for Religious Questions
Farah Atif, Nursultan Askarbekuly, Kareem Darwish +1
Despite the increasing usage of Large Language Models (LLMs) in answering questions in a variety of domains, their reliability and accuracy remain unexamined for a plethora of doma…