21 citations · 122 across the 32 of their papers we have counts for
8 papers · 1 filter
Quintessential inflation and nonlinear effects of the tachyonic trap mechanism
Mindaugas Karčiauskas, Stanislav Rusak, Alejandro Saez
With the help of the tachyonic trapping mechanism one can potentially solve a number of problems affecting quintessential inflation models. In this mechanism we introduce a trappin…
Representation Alignment in Neural Networks
Ehsan Imani, Wei Hu, Martha White
It is now a standard for neural network representations to be trained on large, publicly available datasets, and used for new problems. The reasons for why neural network represent…
Off-Policy Actor-Critic with Emphatic Weightings
Eric Graves, Ehsan Imani, Raksha Kumaraswamy +1
A variety of theoretically-sound policy gradient algorithms exist for the on-policy setting due to the policy gradient theorem, which provides a simplified form for the gradient. T…
Exploiting Action Impact Regularity and Exogenous State Variables for Offline Reinforcement Learning
Vincent Liu, James R. Wright, Martha White
Offline reinforcement learning -- learning a policy from a batch of data -- is known to be hard for general MDPs. These results motivate the need to look at specific classes of MDP…
Greedification Operators for Policy Optimization: Investigating Forward and Reverse KL Divergences
Alan Chan, Hugo Silva, Sungsu Lim +3
Approximate Policy Iteration (API) algorithms alternate between (approximate) policy evaluation and (approximate) greedification. Many different approaches have been explored for a…
Predictive Representation Learning for Language Modeling
Qingfeng Lan, Luke Kumar, Martha White +1
To effectively perform the task of next-word prediction, long short-term memory networks (LSTMs) must keep track of many types of information. Some information is directly related…