agent adaptation 1counterfactual reasoning 1dynamic tool integration 1reinforcement learning 1tool-augmented language models 1
From the 1 of 9 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning
Han Zheng, Yining Ma, Karthick Gunasekaran +4
In LLM Reinforcement Fine-Tuning (RFT), curriculum learning drives both efficiency and performance. Yet, current methods externalize curriculum judgment via handcrafted heuristics…
cs.LG2025
Learning to Segment for Vehicle Routing Problems
Wenbin Ouyang, Sirui Li, Yining Ma +1
Iterative heuristics are widely recognized as state-of-the-art for Vehicle Routing Problems (VRPs). In this work, we exploit a critical observation: a large portion of the solution…