3 citations · 3 across the 6 of their papers we have counts for
1 paper · 2 filters
Hongyu Cai, Arjun Arunasalam, Leo Y. Lin +2
Large language models (LLMs) have become increasingly integrated with various applications. To ensure that LLMs do not generate unsafe responses, they are aligned with safeguards t…