1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2025
L0: Reinforcement Learning to Become General Agents
Junjie Zhang, Jingyi Xi, Zhuoyang Song +7
Training large language models (LLMs) to act as autonomous agents for multi-turn, long-horizon tasks remains significant challenges in scalability and training efficiency. To addre…
hep-ex2024★ 1 cited
Study of the decay and production properties of and
M. Ablikim, M. N. Achasov, P. Adlarson +650
The and processes are studied using data samples collected with the BESIII detector at center-of-m…