agent adaptability 1dynamic tool evolution 1LLM agents 1model context protocol 1tool-use benchmarking 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
MCPEvol-Bench: Benchmarking LLM Agent Performance Across Dynamic Evolutions of MCP Servers
Huanxi Liu, Kun Hu, Jiaqi Liao +6
The paper introduces MCPEvol-Bench, a benchmark that tests how well large language model agents adapt to changing tool interfaces and functionalities in Model Context Protocol (MCP…
cs.SE2024
AutoFeedback: An LLM-based Framework for Efficient and Accurate API Request Generation
Huanxi Liu, Jiaqi Liao, Dawei Feng +2
Large Language Models (LLMs) leverage external tools primarily through generating the API request to enhance task completion efficiency. The accuracy of API request generation sign…