1 paper · 1 filter
Baihui Wang, Bernard Koch
Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophancy as a one-dimensional failure…