1 paper
Wenkai Yang, Shiqi Shen, Guangyao Shen +5
Superalignment, where humans act as weak supervisors for superhuman models, has become a crucial problem with the rapid development of Large Language Models (LLMs). Recent work has…