3 papers
cs.LG2025
Tree-based Dialogue Reinforced Policy Optimization for Red-Teaming Attacks
Ruohao Guo, Afshin Oroojlooy, Roshan Sridhar +3
Despite recent rapid progress in AI safety, current large language models remain vulnerable to adversarial attacks in multi-turn interaction settings, where attackers strategically…
cs.LG2021
Towards optimized actions in critical situations of soccer games with deep reinforcement learning
Pegah Rahimian, Afshin Oroojlooy, Laszlo Toka
Soccer is a sparse rewarding game: any smart or careless action in critical situations can change the result of the match. Therefore players, coaches, and scouts are all curious ab…
cs.LG2020
AttendLight: Universal Attention-Based Reinforcement Learning Model for Traffic Signal Control
Afshin Oroojlooy, Mohammadreza Nazari, Davood Hajinezhad +1
We propose AttendLight, an end-to-end Reinforcement Learning (RL) algorithm for the problem of traffic signal control. Previous approaches for this problem have the shortcoming tha…