agent adaptability 1dynamic tool evolution 1LLM agents 1model context protocol 1tool-use benchmarking 1
From the 1 of 12 linked papers with an AI index.
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
MCPEvol-Bench: Benchmarking LLM Agent Performance Across Dynamic Evolutions of MCP Servers
Huanxi Liu, Kun Hu, Jiaqi Liao +6
The paper introduces MCPEvol-Bench, a benchmark that tests how well large language model agents adapt to changing tool interfaces and functionalities in Model Context Protocol (MCP…
cs.AI2025
Pay More Attention to the Robustness of Prompt for Instruction Data Mining
Qiang Wang, Dawei Feng, Xu Zhang +4
Instruction tuning has emerged as a paramount method for tailoring the behaviors of LLMs. Recent work has unveiled the potential for LLMs to achieve high performance through fine-t…
cs.AI2024
Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models
Yuanzhao Zhai, Tingkai Yang, Kele Xu +4
Agents significantly enhance the capabilities of standalone Large Language Models (LLMs) by perceiving environments, making decisions, and executing actions. However, LLM agents st…