1 paper
Prabhjot Singh, Abhishek Gupta, Chris Betz +4
We reframe clinician overrides of clinical AI recommendations as implicit preference data - the same signal structure exploited by reinforcement learning from human feedback (RLHF)…