3 papers
cs.LG2026
Interpretable GOHR Agents via Sparse Autoencoders
Shiwei Tan, Yusong Zhao, Weiyi Qin +6
A central challenge in interpreting learned decision-making systems is to determine whether their internal representations contain concepts that help explain their behavior. We rep…
cs.RO2026
CoRMA: Contrastive RMA for Contact-Rich Meta-Adaptation
Wentian Wang, Chutong Wen, Hongxu Ma +6
We present CoRMA(Contrastive Robotic Motor Adaptation), a context-based meta-adaptation framework that modifies RMA for force-dominant assembly. CoRMA replaces raw simulator-parame…
cs.LG2025
Toward a Metrology for Artificial Intelligence: Hidden-Rule Environments and Reinforcement Learning
Christo Mathew, Wentian Wang, Jacob Feldman +4
We investigate reinforcement learning in the Game Of Hidden Rules (GOHR) environment, a complex puzzle in which an agent must infer and execute hidden rules to clear a 66 b…