1 paper · 1 filter
Junhao Shi, Qinyuan Cheng, Zhaoye Fei +3
Aligning powerful AI models on tasks that surpass human evaluation capabilities is the central problem of \textbf{superalignment}. To address this problem, weak-to-strong generaliz…