2 papers
cs.AR2026
ArchEval: Measuring AI Agents as Computer Architects
Chenyu Wang, Zishen Wan, Jeffrey Ma +8
Computer architecture has long used benchmarks to make progress measurable. LLM agents create a different measurement problem: success is not merely writing code or tuning paramete…
cs.AR2026
AgentDSE: Reasoning-Augmented Architectural Design Space Exploration
Chenyu Wang, Jiahe Caroline Shi, David Kong +4
Traditional architectural design space exploration (DSE) is highly inefficient, typically requiring tens of thousands of simulator evaluations across various optimization methods.…