2 papers
cs.AI2026
ARB4WM: An Adversarial Robustness Benchmark for World Models in Continuous Control
Junjian Zhang, Hao Tan, Ruonan Li +3
World models are widely used in robotic and agentic engineering control systems due to their ability to learn latent dynamics for planning and decision-making. As these systems are…
cs.CL2026
Replay What Matters: Off-Policy Replay for Efficient LLM Reinforcement Unlearning
Zirui Pang, Chenlong Zhang, Haosheng Tan +3
LLM unlearning has emerged as a cost-effective alternative to full retraining for removing hazardous knowledge from pretrained models while preserving general utility. Recent RL-ba…