2 citations · 2 across the 2 of their papers we have counts for
1 paper · 1 filter
Satoshi Takahashi, Nobuji Kouno, Masaaki Komatsu +1
Artificial Intelligence (AI) safety systems combine character shaping (e.g., Reinforcement Learning from Human Feedback [RLHF], Constitutional AI), which modifies behavioral distri…