2 papers
cs.LG2019
One-element Batch Training by Moving Window
Przemysław Spurek, Szymon Knop, Jacek Tabor +2
Several deep models, esp. the generative, compare the samples from two distributions (e.g. WAE like AutoEncoder models, set-processing deep networks, etc) in their cost functions.…
cs.LG2019
LOSSGRAD: automatic learning rate in gradient descent
Bartosz Wójcik, Łukasz Maziarka, Jacek Tabor
In this paper, we propose a simple, fast and easy to implement algorithm LOSSGRAD (locally optimal step-size in gradient descent), which automatically modifies the step-size in gra…