2 papers
cs.AI2026
InjectRBP: Steering Large Language Model Reasoning Behavior via Pattern Injection
Xiuping Wu, Zhao Yu, Yuxin Cheng +4
Reasoning can significantly enhance the performance of Large Language Models. While recent studies have exploited behavior-related prompts adjustment to enhance reasoning, these de…
cs.LG2025
Test-driven Reinforcement Learning in Continuous Control
Zhao Yu, Xiuping Wu, Liangjun Ke
Reinforcement learning (RL) has been recognized as a powerful tool for robot control tasks. RL typically employs reward functions to define task objectives and guide agent learning…