3 papers
cs.CY2026
How to Catch a GPU: A Taxonomy of Verification and Enforcement Mechanisms for International AI Agreements
Raymond Koopmanschap, Otto Barten
Several international agreements have been proposed to regulate frontier AI development in response to catastrophic risks. However, there is no structured way to evaluate whether t…
cs.CY2025
Toward a Global Regime for Compute Governance: Building the Pause Button
Ananthi Al Ramiah, Raymond Koopmanschap, Josh Thorsteinson +5
As AI capabilities rapidly advance, the risk of catastrophic harm from large-scale training runs is growing. Yet the compute infrastructure that enables such development remains la…
cs.AI2024
Inducing Human-like Biases in Moral Reasoning Language Models
Artem Karpov, Seong Hah Cho, Austin Meek +3
In this work, we study the alignment (BrainScore) of large language models (LLMs) fine-tuned for moral reasoning on behavioral data and/or brain data of humans performing the same…