agent evaluation 1agent harnesses 1behavior localization 1code analysis 1dense rewards 1llm-assisted tooling 1long-horizon planning 1multi-step tasks 1progressive disclosure 1software engineering 1terminal benchmarks 1
From the 2 of 10 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
RCP-Merging: Merging Long Chain-of-Thought Models with Domain-Specific Models by Considering Reasoning Capability as Prior
Junyao Yang, Jianwei Wang, Huiping Zhuang +2
Large Language Models (LLMs) with long chain-of-thought (CoT) capability, termed Reasoning Models, demonstrate superior intricate problem-solving abilities through multi-step long…
cs.CL2026
ReasonAny: Incorporating Reasoning Capability to Any Model via Simple and Effective Model Merging
Junyao Yang, Chen Qian, Dongrui Liu +3
Large Reasoning Models (LRMs) with long chain-of-thought reasoning have recently achieved remarkable success. Yet, equipping domain-specialized models with such reasoning capabilit…