2 papers
cs.LG2026
ERR+: Sequential Entropy Resolution for Efficient and Decisive LLM Reasoning
Xin Jiang, Minhao Wang, Wen Wu +4
Large reasoning models achieve strong performance on complex tasks by generating extended chain-of-thought (CoT) traces via reinforcement learning with verifiable rewards (RLVR). W…
cs.AI2026
SkillSmith: Enhancing Locally Deployed Agents via Automatic Skill Construction and Evolution
Xinle Jiang, Remy Xie, Ming Tang
LLM-based agent frameworks now act as personal assistants for multi-step tasks. Existing agent frameworks such as OpenClaw commonly follow the Cloud Agent depolyment mode using clo…