1 citations · 2 across the 3 of their papers we have counts for
3 papers
DIP-RL: Demonstration-Inferred Preference Learning in Minecraft
Ellen Novoseller, Vinicius G. Goecks, David Watkins +2
In machine learning for sequential decision-making, an algorithmic agent learns to interact with an environment while receiving feedback in the form of a reward signal. However, in…
Towards Solving Fuzzy Tasks with Human Feedback: A Retrospective of the MineRL BASALT 2022 Competition
Stephanie Milani, Anssi Kanervisto, Karolis Ramanauskas +27
To facilitate research in the direction of fine-tuning foundation models from human feedback, we held the MineRL BASALT Competition on Fine-Tuning from Human Feedback at NeurIPS 20…
Teleoperated Robot Grasping in Virtual Reality Spaces
Jiaheng Hu, David Watkins, Peter Allen
Despite recent advancement in virtual reality technology, teleoperating a high DoF robot to complete dexterous tasks in cluttered scenes remains difficult. In this work, we propose…