3 citations · 3 across the 4 of their papers we have counts for
3 papers
A Systematic Investigation of RL-Jailbreaking in LLMs
Montaser Mohammedalamen, Kevin Roice, Reginald McLean +1
The evolution of generative models from next-token predictors to autonomous engines of complex systems necessitates rigorous safety hardening. Adversarial jailbreaking, the strateg…
Generalization in Monitored Markov Decision Processes (Mon-MDPs)
Montaser Mohammedalamen, Michael Bowling
Reinforcement learning (RL) typically models the interaction between the agent and environment as a Markov decision process (MDP), where the rewards that guide the agent's behavior…
Transfer Learning for Prosthetics Using Imitation Learning
Montaser Mohammedalamen, Waleed D. Khamies, Benjamin Rosman
In this paper, We Apply Reinforcement learning (RL) techniques to train a realistic biomechanical model to work with different people and on different walking environments. We benc…