3 papers
cs.CL2026
Enhancing Agentic RL with Progressive Reward Shaping and Value-based Sampling Policy Optimization
Jianghao Su, Xia Zeng, Luhui Liu +3
Large Language Models (LLMs) empowered with Tool-Integrated Reasoning (TIR) can iteratively plan, call external tools, and integrate returned information to solve complex, long-hor…
cs.SD2026
IKFST: IOO and KOO Algorithms for Accelerated and Precise WFST-based End-to-End Automatic Speech Recognition
Zhuoran Zhuang, Ye Chen, Chao Luo +5
End-to-end automatic speech recognition has become the dominant paradigm in both academia and industry. To enhance recognition performance, the Weighted Finite-State Transducer (WF…
cs.CL2025
CoDA: A Context-Decoupled Hierarchical Agent with Reinforcement Learning
Xuanzhang Liu, Jianglun Feng, Zhuoran Zhuang +7
Large Language Model (LLM) agents trained with reinforcement learning (RL) show great promise for solving complex, multi-step tasks. However, their performance is often crippled by…