1 citations · 1 across the 6 of their papers we have counts for
Showing cs.CYShow all
2 papers · 1 filter
cs.CY2026
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
Susanne Gaube, Markus Langer, Tim Miller +17
The use of Artificial Intelligence (AI) in high-risk, decision-making scenarios presents technical, safety, and normative challenges; problems that may only be ameliorated by human…
cs.CY2026
The Story is Not the Science: Execution-Grounded Evaluation of Mechanistic Interpretability Research
Xiaoyan Bai, Alexander Baumgartner, Haojia Sun +2
Reproducibility crises across sciences highlight the limitations of the paper-centric review system in assessing the rigor and reproducibility of research. AI agents that autonomou…