1 paper
Xiaolong Jin, Zhuo Zhang, Xiangyu Zhang
Large Language Model (LLM) alignment aims to ensure that LLM outputs match with human values. Researchers have demonstrated the severity of alignment problems with a large spectrum…