2 papers
cs.CL2026
Evaluation Awareness in Language Models: Representation, Verbalization, and Control
Farzaneh Heidari, Amin Memarian, Guillaume Rabusseau
Both capability and safety benchmarks rest upon the assumption that the behavior of language models undergoing a test is informative about their behavior in deployment. This assump…
cs.AI2026
Position: Collusion Risks Among AI Reasoning Agents Justify Certification Requirements for Making Market Decisions
Matthew Riemer, Tommaso Tosato, Amin Memarian +4
This position paper argues that AI agents with chain-of-thought reasoning capabilities are predisposed to exhibit collusive behavior and should be required to obtain behavioral cer…