Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
Zhongyi Li, Wan Tian, Jinju Chen +4
Collaborative multi-agent large language models (LLMs) can solve complex reasoning tasks by decomposing roles, but reinforcement learning for such systems is limited by credit assi…
cs.AI2026
InA-Probe: Instruction-Aware Active Probing for Time Series Forecasting with LLMs
Peiliang Gong, Emadeldeen Eldele, Chenyu Liu +8
Large Language Models (LLMs) have recently demonstrated impressive potential for time series forecasting. However, existing methods predominantly rely on passive modality alignment…
cs.AI2025
ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection
Jiaqi Li, Xinyi Dong, Yang Liu +6
We present a novel pipeline, ReflectEvo, to demonstrate that small language models (SLMs) can enhance meta introspection through reflection learning. This process iteratively gener…