2 papers
cs.LG2024
On Designing Effective RL Reward at Training Time for LLM Reasoning
Jiaxuan Gao, Shusheng Xu, Wenjie Ye +6
Reward models have been increasingly critical for improving the reasoning capability of LLMs. Existing research has shown that a well-trained reward model can substantially improve…
physics.atom-ph2024
Evaluation of the systematic error induced by quadratic Zeeman effect using hyperfine ground state exchange method in a long-baseline dual-species atom interferometer
Yu-Hang Ji, Chuan He, Si-Tong Yan +7
The systematic error induced by the quadratic Zeeman effect is non-negligible in atom interferometers and must be precisely evaluated. We theoretically analyze the phase shift indu…