papers
Publications (2)
cs.LG2022
Explaining Reinforcement Learning Policies through Counterfactual Trajectories
Julius Frost, Olivia Watkins, Eric Weiner +4
In order for humans to confidently decide where to employ RL agents for real-world tasks, a human developer must validate that the agent will perform well at test-time. Some policy…
cs.LG2022
Neural Parameter Allocation Search
Bryan A. Plummer, Nikoli Dryden, Julius Frost +2
Training neural networks requires increasing amounts of memory. Parameter sharing can reduce memory and communication costs, but existing methods assume networks have many identica…