3 papers
cs.AI2026
Evaluating the Search Agent in a Parallel World
Jiawei Chen, Xintian Shen, Lihao Zheng +7
Integrating web search tools has significantly extended the capability of LLMs to address open-world, real-time, and long-tail problems. However, evaluating these Search Agents pre…
cs.CL2026
Evaluating LLMs' Divergent Thinking Capabilities for Scientific Idea Generation with Minimal Context
Kai Ruan, Xuan Wang, Jixiang Hong +3
While Large Language Models (LLMs) demonstrate remarkable capabilities in scientific tasks such as literature analysis and experimental design (e.g., accurately extracting key find…
cs.MA2025
Benchmarking LLMs' Swarm intelligence
Kai Ruan, Mowen Huang, Ji-Rong Wen +1
Large Language Models (LLMs) show potential for complex reasoning, yet their capacity for emergent coordination in Multi-Agent Systems (MAS) when operating under strict swarm-like…