3 papers
cs.LG2020
Stabilizing Transformer-Based Action Sequence Generation For Q-Learning
Gideon Stein, Andrey Filchenkov, Arip Asadulaev
Since the publication of the original Transformer architecture (Vaswani et al. 2017), Transformers revolutionized the field of Natural Language Processing. This, mainly due to thei…
cs.LG2019
Conditioning of Reinforcement Learning Agents and its Policy Regularization Application
Arip Asadulaev, Igor Kuznetsov, Gideon Stein +1
The outcome of Jacobian singular values regularization was studied for supervised learning problems. It also was shown that Jacobian conditioning regularization can help to avoid t…
cs.LG2019
Interpretable Few-Shot Learning via Linear Distillation
Arip Asadulaev, Igor Kuznetsov, Andrey Filchenkov
It is important to develop mathematically tractable models than can interpret knowledge extracted from the data and provide reasonable predictions. In this paper, we present a Line…