1 paper
Gabriel Recchia, Chatrik Singh Mangat, Jinu Nyachhyon +4
Scalable oversight protocols aim to empower evaluators to accurately verify AI models more capable than themselves. However, human evaluators are subject to biases that can lead to…