2 citations · 2 across the 1 of their papers we have counts for
1 paper
Wenqi Marshall Guo, Yiyang Du, Heidi J. S. Tworek +1
Large Language Models (LLMs) are usually aligned with "human values/preferences" to prevent harmful output. Discussions around the alignment of Large Language Models (LLMs) general…