1 paper
Abhimanyu Pallavi Sudhir, Jackson Kaunismaa, Arjun Panickssery
As AI agents surpass human capabilities, scalable oversight -- the problem of effectively supplying human feedback to potentially superhuman AI models -- becomes increasingly criti…