collaborators

5 papers

cs.AI2026

When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines

Yiyao Zhang, Diksha Goel, Hussain Ahmad +2

An LLM judge deployed inside a reasoning pipeline does not merely measure quality, it decides which answer ships. We show that the cost of that decision depends less on judge accur…

cs.AI2026

CausalNav: Reliability-Certified Causal World Models for Control under Physical-Parameter Shift

Yiyao Zhang, Diksha Goel, Hussain Ahmad +2

A world model is only useful for physical AI if it changes what the agent does, and only safe if it declines to do so when it is wrong. We study both halves of that requirement wit…

cs.CR2026

AgenticVM: Agentic AI for Adaptive Software Vulnerability Management

Asrul Arifin, Hussain Ahmad, Yiyao Zhang +1

As software systems grow in scale and complexity, vulnerability management is increasingly strained by high alert volumes, fragmented toolchains, and manual triage processes. We in…

cs.CR2026

Explainable Autonomous Cyber Defense using Adversarial Multi-Agent Reinforcement Learning

Yiyao Zhang, Diksha Goel, Hussain Ahmad

Autonomous agents are increasingly deployed in both offensive and defensive cyber operations, creating high-speed, closed-loop interactions in critical infrastructure environments.…

q-fin.PM2025

RegimeFolio: A Regime Aware ML System for Sectoral Portfolio Optimization in Dynamic Markets

Yiyao Zhang, Diksha Goel, Hussain Ahmad +1

Financial markets are inherently non-stationary, with shifting volatility regimes that alter asset co-movements and return distributions. Standard portfolio optimization methods, t…