73 citations · 73 across the 1 of their papers we have counts for
1 paper · 1 filter
Yang Liu, Yuanshun Yao, Jean-Francois Ton +6
Ensuring alignment, which refers to making models behave in accordance with human intentions [1,2], has become a critical task before deploying large language models (LLMs) in real…