1 citations · 1 across the 3 of their papers we have counts for
3 papers
Towards Solving Fuzzy Tasks with Human Feedback: A Retrospective of the MineRL BASALT 2022 Competition
Stephanie Milani, Anssi Kanervisto, Karolis Ramanauskas +27
To facilitate research in the direction of fine-tuning foundation models from human feedback, we held the MineRL BASALT Competition on Fine-Tuning from Human Feedback at NeurIPS 20…
Improving performance in multi-objective decision-making in Bottles environments with soft maximin approaches
Benjamin J Smith, Robert Klassert, Roland Pihlakas
Balancing multiple competing and conflicting objectives is an essential task for any artificial intelligence tasked with satisfying human values or preferences. Conflict arises bot…
Variational learning of quantum ground states on spiking neuromorphic hardware
Robert Klassert, Andreas Baumbach, Mihai A. Petrovici +1
Recent research has demonstrated the usefulness of neural networks as variational ansatz functions for quantum many-body states. However, high-dimensional sampling spaces and trans…