activity
20162022
most citedAPPLE: Adaptive Planner Parameter Learning from Evaluative Feedback

34 citations · 59 across the 10 of their papers we have counts for

collaborators
Showing cs.AIShow all

6 papers · 1 filter

cs.AI202114 cited

Recent Advances in Leveraging Human Guidance for Sequential Decision-Making Tasks

Ruohan Zhang, Faraz Torabi, Garrett Warnell +1

A longstanding goal of artificial intelligence is to create artificial agents capable of learning to perform tasks that require sequential decision making. Importantly, while it is…

cs.AI2020

An Imitation from Observation Approach to Transfer Learning with Dynamics Mismatch

Siddharth Desai, Ishan Durugkar, Haresh Karnan +3

We examine the problem of transferring a policy learned in a source environment to a target environment with different dynamics, particularly in the case where it is critical to re…

cs.AI20194 cited

A Narration-based Reward Shaping Approach using Grounded Natural Language Commands

Nicholas Waytowich, Sean L. Barton, Vernon Lawhern +1

While deep reinforcement learning techniques have led to agents that are successfully able to learn to perform a number of tasks that had been previously unlearnable, these techniq…

cs.AI2018

Deterministic Implementations for Reproducibility in Deep Reinforcement Learning

Prabhat Nagarajan, Garrett Warnell, Peter Stone

While deep reinforcement learning (DRL) has led to numerous successes in recent years, reproducing these successes can be extremely challenging. One reproducibility challenge parti…

cs.AI2018

Behavioral Cloning from Observation

Faraz Torabi, Garrett Warnell, Peter Stone

Humans often learn how to perform tasks via imitation: they observe others perform a task, and then very quickly infer the appropriate actions to take based on their observations.…

cs.AI2017

Deep TAMER: Interactive Agent Shaping in High-Dimensional State Spaces

Garrett Warnell, Nicholas Waytowich, Vernon Lawhern +1

While recent advances in deep reinforcement learning have allowed autonomous learning agents to succeed at a variety of complex tasks, existing algorithms generally require a lot o…