14 citations · 17 across the 6 of their papers we have counts for
1 paper · 2 filters
Trevor Ablett, Bryan Chan, Jayce Haoran Wang +1
Common approaches to providing feedback in reinforcement learning are the use of hand-crafted rewards or full-trajectory expert demonstrations. Alternatively, one can use examples…