4 citations · 4 across the 4 of their papers we have counts for
4 papers
Steering Large Language Model Activations in Sparse Spaces
Reza Bayat, Ali Rahimi-Kalahroudi, Mohammad Pezeshki +2
A key challenge in AI alignment is guiding large language models (LLMs) to follow desired behaviors at test time. Activation steering, which modifies internal model activations dur…
Partial Models for Building Adaptive Model-Based Reinforcement Learning Agents
Safa Alver, Ali Rahimi-Kalahroudi, Doina Precup
In neuroscience, one of the key behavioral tests for determining whether a subject of study exhibits model-based behavior is to study its adaptiveness to local changes in the envir…
Replay Buffer with Local Forgetting for Adapting to Local Environment Changes in Deep Model-Based Reinforcement Learning
Ali Rahimi-Kalahroudi, Janarthanan Rajendran, Ida Momennejad +2
One of the key behavioral characteristics used in neuroscience to determine whether the subject of study -- be it a rodent or a human -- exhibits model-based learning is effective…
Towards Evaluating Adaptivity of Model-Based Reinforcement Learning Methods
Yi Wan, Ali Rahimi-Kalahroudi, Janarthanan Rajendran +3
In recent years, a growing number of deep model-based reinforcement learning (RL) methods have been introduced. The interest in deep model-based RL is not surprising, given its man…