#long-horizon reasoning
4 papers · 1 filter
ConMem: Contribution-Aware Memory for Long-Horizon Manufacturing Inspection Logs
Bingchen Liu, Yuanyuan Fang, Lei Liu +6
ConMem is a memory framework that selects and retains the most diagnostically valuable segments of long‑horizon manufacturing inspection logs for LLM‑based analysis, using Shapley‑…
Contrastive Reinforced Policy Optimization via Privileged Self-Distillation
Xingjian Wu, Junlin Liu, Xingchen Liu +6
The paper introduces Contrastive Reinforced Policy Optimization (CRPO), a method that frames on‑policy self‑distillation for large language models as a contrastive learning problem…
A hierarchical memory architecture overcomes context limits in long-horizon multi-agent computational modeling
Shivendra G. Tewari, Holly Kimko
The paper introduces Ensemble QSP, a multi‑agent framework with a three‑layer hierarchical memory that keeps context size bounded, enabling long‑horizon autonomous modeling of phar…
Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent
Lei Bai, Zongsheng Cao, Yang Chen +50
The paper introduces Agents-A1, a 35B mixture-of-experts agent model that attains trillion-parameter-level performance by extending the length of reasoning horizons and integrating…