2 papers
cs.LG2026
StarOR: Synergizing Tree Search and Test-Time Reinforcement Learning for Optimization Modeling
Jiajun Li, Yu Ding, Shisi Guan +2
Optimization modeling is inherently hierarchical, requiring a precise sequence of symbolic commitments. Traditional learning-based automated optimization modeling methods improve m…
cs.CL2026
SkillWiki: A Living Knowledge Infrastructure for Agent Skills
Dingcheng Huang, Yuda Ding, Bingshuo Liu +8
While knowledge is managed through Wikipedia and software through GitHub, agent skills still lack an infrastructure for large-scale production, governance, and evolution. SkillWiki…