#long-horizon reasoning

topiclong-horizon reasoning

4 papers · 1 filter

cs.AI2026

ConMem: Contribution-Aware Memory for Long-Horizon Manufacturing Inspection Logs

Bingchen Liu, Yuanyuan Fang, Lei Liu +6

ConMem is a memory framework that selects and retains the most diagnostically valuable segments of long‑horizon manufacturing inspection logs for LLM‑based analysis, using Shapley‑…

cs.LG2026

Contrastive Reinforced Policy Optimization via Privileged Self-Distillation

Xingjian Wu, Junlin Liu, Xingchen Liu +6

The paper introduces Contrastive Reinforced Policy Optimization (CRPO), a method that frames on‑policy self‑distillation for large language models as a contrastive learning problem…

q-bio.QM2026

A hierarchical memory architecture overcomes context limits in long-horizon multi-agent computational modeling

Shivendra G. Tewari, Holly Kimko

The paper introduces Ensemble QSP, a multi‑agent framework with a three‑layer hierarchical memory that keeps context size bounded, enabling long‑horizon autonomous modeling of phar…

cs.CL2026

Scaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B Agent

Lei Bai, Zongsheng Cao, Yang Chen +50

The paper introduces Agents-A1, a 35B mixture-of-experts agent model that attains trillion-parameter-level performance by extending the length of reasoning horizons and integrating…