2 papers
cs.AI2026
Beyond Instruction Following: Learning Grounded Skill-Following with Skill Contracts
Jianghan Shen, Zhenjie Liu, Yue Li +10
Instruction following typically enforces discrete, response-level requirements, whereas an expert-authored skill prescribes procedural requirements spanning multiple phases and env…
cs.AI2026
CuSearch: Curriculum Rollout Sampling via Search Depth for Agentic RAG
Jianghan Shen, Siqi Luo, Xinyu Cheng +6
Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for training agentic retrieval-augmented generation (RAG) systems from outcome-only superv…