From the 1 of 2 linked papers with an AI index.
2 papers
cs.LG2026
TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling
Yanyu Ren, Xizheng Wang, Xiao Liu +8
Reinforcement learning (RL) for large language models is moving toward multi-turn agentic workloads, where rollout tasks repeatedly pause for external environments, resume with gro…
cs.AI2026
An Empirical Study of Coordination Mode as the First-Class Citizen in From-Scratch Multi-Agent Coding
Yanyu Ren, Yunfeng Bai, Xizheng Wang +2
The paper presents MSEval, a benchmark that evaluates how multi‑agent coding systems build real‑world software, measuring functional success, latency, and token cost while varying…