works on

From the 1 of 6 linked papers with an AI index.

collaborators

6 papers

cs.LG2026

DACRI: Decision-Aware Causal Intervention Ranking for Critical Supply Chains

Shiqi Huang, Jiani He, Dingyan Shang +4

Detecting or attributing a supply-chain disruption is not the same as selecting the intervention that maximizes recoverable net value. We present CriticalSCM-Bench v1, a controlled…

cs.AI2026

Safety, or Just Capability? A Validity Audit of Agent-Safety Benchmarks

Youting Wang, Xiao Han, Dingyan Shang +2

Agent-safety benchmarks measure different behaviors, and their scores get quoted interchangeably as an agent's safety. We treat four of them (R-Judge, InjecAgent, AgentHarm, AgentD…

cs.LG2026

Accuracy-Preserving Stability Regularization for Large-Scale Retail Demand Forecasting

Jize Li, Jiani He, Dishu Yang +3

The paper proposes a training-time regularization method that penalizes abrupt changes between consecutive forecasts to improve stability while keeping point accuracy nearly unchan…

cs.AI2026

Context-Masked Truncated Reasoning Audits for Answer-Key Dependence in LLM Tutors

Bonan Shen, Dingyan Shang, Youting Wang +2

Large language model (LLM) tutors may have access to teacher notes, answer keys, rubrics, or retrieved solutions while producing student-facing explanations. We study whether trunc…

cs.AI2026

Self-Commitment Latency: A Reward-Free Probe for Prompted Implicit Hacking

Bonan Shen, Youting Wang, Dingyan Shang +1

Implicit reward hacking is hard to audit when a language model's chain of thought appears benign: a final answer may be anchored by a prompt shortcut while the written reasoning st…

cs.LG2026

When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL

Youting Wang, Yuan Tang, Bowen Liu +2

For sparse, structured reinforcement-learning tasks with semantic reward-function interfaces, LLM-generated reward shaping is better framed as debugging than one-shot generation. W…