1 paper
Benjamin Arnav, Pablo Bernabeu-Pérez, Nathan Helm-Burger +3
As AI models are deployed with increasing autonomy, it is important to ensure they do not take harmful actions unnoticed. As a potential mitigation, we investigate Chain-of-Thought…