2 papers
cs.AI2026
SMaRT: Select, Mix, and ReinvenT -- A Strategy Fusion Framework for LLM-Driven Reasoning and Planning
Nikhil Verma, Manasa Bharadwaj, Wonjun Jang +4
Large Language Models (LLMs) have redefined complex task automation with exceptional generalization capabilities. Despite these advancements, state-of-the-art methods rely on singl…
cs.CL2025
GEMMAS: Graph-based Evaluation Metrics for Multi Agent Systems
Jisoo Lee, Raeyoung Chang, Dongwook Kwon +2
Multi-agent systems built on language models have shown strong performance on collaborative reasoning tasks. However, existing evaluations focus only on the correctness of the fina…