35 citations · 95 across the 9 of their papers we have counts for
1 paper · 2 filters
Xu He, Haipeng Chen, Bo An
Human feedback is widely used to train agents in many domains. However, previous works rarely consider the uncertainty when humans provide feedback, especially in cases that the op…