1 paper · 1 filter
Xufei Lv, Kehai Chen, Haoyuan Sun +3
Alignment of large language models (LLMs) with human values has recently garnered significant attention, with prominent examples including the canonical yet costly Reinforcement Le…