1 paper · 1 filter
Kush R. Varshney, Zahra Ashktorab, Djallel Bouneffouf +2
Much of the research focus on AI alignment seeks to align large language models and other foundation models to the context-less and generic values of helpfulness, harmlessness, and…