6 citations · 6 across the 3 of their papers we have counts for
3 papers
cs.IR2025
Evaluation Report on MCP Servers
Zhiling Luo, Xiaorong Shi, Xuanrui Lin +1
With the rise of LLMs, a large number of Model Context Protocol (MCP) services have emerged since the end of 2024. However, the effectiveness and efficiency of MCP servers have not…
cs.CR2025
SecBench: A Comprehensive Multi-Dimensional Benchmarking Dataset for LLMs in Cybersecurity
Pengfei Jing, Mengyun Tang, Xiaorong Shi +5
Evaluating Large Language Models (LLMs) is crucial for understanding their capabilities and limitations across various applications, including natural language processing and code…
cs.AI2024★ 6 cited
A Preview of XiYan-SQL: A Multi-Generator Ensemble Framework for Text-to-SQL
Yingqi Gao, Yifu Liu, Xiaoxia Li +10
To tackle the challenges of large language model performance in natural language to SQL tasks, we introduce XiYan-SQL, an innovative framework that employs a multi-generator ensemb…