1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2021
Can Q-learning solve Multi Armed Bantids?
Refael Vivanti
When a reinforcement learning (RL) method has to decide between several optional policies by solely looking at the received reward, it has to implicitly optimize a Multi-Armed-Band…
cs.LG2019★ 1 cited
Adaptive Symmetric Reward Noising for Reinforcement Learning
Refael Vivanti, Talya D. Sohlberg-Baris, Shlomo Cohen +1
Recent reinforcement learning algorithms, though achieving impressive results in various fields, suffer from brittle training effects such as regression in results and high sensiti…