14 citations · 21 across the 5 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2022★ 5 cited
Improving Policy Learning via Language Dynamics Distillation
Victor Zhong, Jesse Mu, Luke Zettlemoyer +2
Recent work has shown that augmenting environments with language descriptions improves policy learning. However, for environments with complex language abstractions, learning how t…
cs.LG2022★ 14 cited
Active Learning Helps Pretrained Models Learn the Intended Task
Alex Tamkin, Dat Nguyen, Salil Deshpande +2
Models can fail in unpredictable ways during deployment due to task ambiguity, when multiple behaviors are consistent with the provided training data. An example is an object class…
cs.LG2020
Compositional Explanations of Neurons
Jesse Mu, Jacob Andreas
We describe a procedure for explaining neurons in deep representations by identifying compositional logical concepts that closely approximate neuron behavior. Compared to prior wor…