1 paper · 1 filter
Yisu Wang, Ming Wang, Haoyuan Song +4
Post-training for large language models (LLMs) is constrained by the high cost of acquiring new knowledge or correcting errors and by the unintended side effects that frequently ar…