1 paper · 1 filter
Yurui Pan, Ke Xu, Bo Peng
Alignment of large language models (LLMs) via SFT and RLHF/DPO typically ignores the global geometry of the representation space, relying instead on local token likelihoods or scal…