2 papers
cs.AI2025
Evaluation Awareness Scales Predictably in Open-Weights Large Language Models
Maheep Chaudhary, Ian Su, Nikhil Hooda +6
Large language models (LLMs) can internally distinguish between evaluation and deployment contexts, a behaviour known as \emph{evaluation awareness}. This undermines AI safety eval…
cs.LG2022
Renaissance Robot: Optimal Transport Policy Fusion for Learning Diverse Skills
Julia Tan, Ransalu Senanayake, Fabio Ramos
Deep reinforcement learning (RL) is a promising approach to solving complex robotics problems. However, the process of learning through trial-and-error interactions is often highly…