12 citations · 13 across the 2 of their papers we have counts for
1 paper · 1 filter
Luis Haug, Sebastian Tschiatschek, Adish Singla
Learning near-optimal behaviour from an expert's demonstrations typically relies on the assumption that the learner knows the features that the true reward function depends on. In…