1 citations · 1 across the 11 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
CHIME: Credit-Aware Hierarchical Memory Evolution for Long-Horizon Agentic Planning
Yongshi Ye, Tian Lan, Feihu Jiang +7
Planning is a central capability that enables agents to decompose complex long-horizon tasks into manageable steps. Test-time search and training-based methods improve planning but…
cs.AI2026
TARL: Transaction-Aware Reliable Ledgers for Executable Memory Management in Long-Term Agents
Han Xiao, Hongjun Xu, Xin Zhang +2
Persistent memory helps long-term agents retain knowledge, yet a single update error can repeatedly distort future retrieval and reasoning. Most existing systems reduce memory upda…
cs.AI2026
Don't Peek at the Answer: Outcome-Masked Group Relative Policy Optimization for Label-Free RLVR
Yongshi Ye, Liang Zhang, Yidong Chen +2
Reinforcement Learning with Verifiable Rewards (RLVR) improves LLM reasoning but typically relies on ground-truth (GT) answers, limiting scalability. Voting-based label-free RLVR r…