1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2025★ 1 cited
Agentic Reinforcement Learning with Implicit Step Rewards
Xiaoqian Liu, Ke Wang, Yuchuan Wu +4
Large language models (LLMs) are increasingly developed as autonomous agents using reinforcement learning (agentic RL) that reason and act in interactive environments. However, spa…
astro-ph.SR2025
Characterization and formation of the Mg i 12.32 μm line in the quiet Sun and sunspot
Yuchuan Wu, Wenxian Li, Xianyong Bai +3
The Mg I 12.32 μm line is highly sensitive to magnetic fields due to its long wavelength, making it a promising tool for precise solar-magnetic-field measurements. The formation of…
cs.CL2025
TimeHC-RL: Temporal-aware Hierarchical Cognitive Reinforcement Learning for Enhancing LLMs' Social Intelligence
Guiyang Hou, Xing Gao, Yuchuan Wu +8
Recently, Large Language Models (LLMs) have made significant progress in IQ-related domains that require careful thinking, such as mathematics and coding. However, enhancing LLMs'…