2 papers
cs.SE2026
Coding Agents as Test-Suite Auditors: Finding What Official Suites Miss While Approaching What They Catch
Shuyang Xie, Shuxiao Xie, Feng Zhu +2
Online-judge verdicts and the datasets and benchmarks built on them are treated as ground truth for evaluating and training large language models for code. Yet prior audits have so…
cs.CV2026
OPERA: An Agent for Image Restoration with End-to-End Joint Planning-Execution Optimization
Feng Zhu, Shuyang Xie, Yihan Zeng +2
Real-world image restoration is challenging due to complex and interacting mixed degradations. Recent agent-based approaches address this problem by composing multiple task-specifi…