21 citations · 33 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 21 cited
Reward Design with Language Models
Minae Kwon, Sang Michael Xie, Kalesha Bullard +1
Reward design in reinforcement learning (RL) is challenging since specifying human notions of desired behavior may be difficult via reward functions or require many expert demonstr…
cs.LG2021★ 12 cited
Extending the WILDS Benchmark for Unsupervised Adaptation
Shiori Sagawa, Pang Wei Koh, Tony Lee +17
Machine learning systems deployed in the wild are often trained on a source distribution but deployed on a different target distribution. Unlabeled data can be a powerful point of…