1 citations · 1 across the 1 of their papers we have counts for
1 paper
Maarten Buyl, Hadi Khalaf, Claudio Mayrink Verdun +3
In AI alignment, extensive latitude must be granted to annotators, either human or algorithmic, to judge which model outputs are `better' or `safer.' We refer to this latitude as a…