behavioral tokens 1exploration 1reinforcement learning 1robotic manipulation 1sample efficiency 1vision-language-action 1
From the 1 of 6 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
QPO: Query-dependent Prompt Optimization via Multi-Loop Offline Reinforcement Learning
Yilun Kong, Hangyu Mao, Qi Zhao +7
Prompt engineering has demonstrated remarkable success in enhancing the performance of large language models (LLMs) across diverse tasks. However, most existing prompt optimization…
cs.AI2025
Neuron-level Balance between Stability and Plasticity in Deep Reinforcement Learning
Jiahua Lan, Sen Zhang, Haixia Pan +3
In contrast to the human ability to continuously acquire knowledge, agents struggle with the stability-plasticity dilemma in deep reinforcement learning (DRL), which refers to the…