2 citations · 3 across the 5 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.LG2024★ 1 cited
Improving Actor-Critic Training with Steerable Action-Value Approximation Errors
Bahareh Tasdighi, Nicklas Werge, Yi-Shan Wu +1
Off-policy actor-critic algorithms have shown strong potential in deep reinforcement learning for continuous control tasks. Their success primarily comes from leveraging pessimisti…
cs.LG2024
Deep Exploration with PAC-Bayes
Bahareh Tasdighi, Manuel Haussmann, Nicklas Werge +2
Reinforcement learning (RL) for continuous control under delayed rewards is an under-explored problem despite its significance in real-world applications. Many complex skills are b…