Showing 2025Show all
2 papers · 1 filter
quant-ph2025
Group-Theoretic Reinforcement Learning of Dynamical Decoupling Sequences
Charles Marrder, Shuo Sun, Murray J. Holland
Dynamical decoupling seeks to mitigate phase decoherence in qubits by applying a carefully designed sequence of effectively instantaneous electromagnetic pulses. Although analytic…
cs.LG2025
Generalization of RLVR Using Causal Reasoning as a Testbed
Brian Lu, Hongyu Zhao, Shuo Sun +3
Reinforcement learning with verifiable rewards (RLVR) has emerged as a promising paradigm for post-training large language models (LLMs) on complex reasoning tasks. Yet, the condit…