65 citations · 124 across the 8 of their papers we have counts for
Showing 2018Show all
2 papers · 1 filter
cs.LG2018
Deep Variational Reinforcement Learning for POMDPs
Maximilian Igl, Luisa Zintgraf, Tuan Anh Le +2
Many real-world sequential decision making problems are partially observable by nature, and the environment model is typically unknown. Consequently, there is great need for reinfo…
stat.ML2018
Tighter Variational Bounds are Not Necessarily Better
Tom Rainforth, Adam R. Kosiorek, Tuan Anh Le +4
We provide theoretical and empirical evidence that using tighter evidence lower bounds (ELBOs) can be detrimental to the process of learning an inference network by reducing the si…