3 papers
cs.RO2026
GPU-Parallel Multi-Task Reinforcement Learning with Demonstration Guided Policy Optimization
Rui Zhang, Qiwei Wu, Zhengyu Zhang +5
Large scale GPU-parallel reinforcement learning has changed what can be trained in robot simulation, yet most systems still optimize one specialist policy per task. We propose a co…
cs.CR2025
SCOUT: A Defense Against Data Poisoning Attacks in Fine-Tuned Language Models
Mohamed Afane, Abhishek Satyam, Ke Chen +3
Backdoor attacks create significant security threats to language models by embedding hidden triggers that manipulate model behavior during inference, presenting critical risks for…
cs.LG2025
Stackelberg Coupling of Online Representation Learning and Reinforcement Learning
Fernando Martinez, Tao Li, Yingdong Lu +1
Deep Q-learning jointly learns representations and values within monolithic networks, promising beneficial co-adaptation between features and value estimates. Although this archite…