Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
ProtocolBench: Which LLM MultiAgent Protocol to Choose?
Hongyi Du, Jiaqi Su, Jisen Li +6
As large-scale multi-agent systems evolve, the communication protocol layer has become a critical yet under-evaluated factor shaping performance and reliability. Despite the existe…
cs.AI2025
ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
Cheng Qian, Hongyi Du, Hongru Wang +6
Recent progress in large language models (LLMs) has enabled substantial advances in solving mathematical problems. However, existing benchmarks often fail to reflect the complexity…