From the 1 of 13 linked papers with an AI index.
13 papers
One Run Is Not an Idea: The Implementation Lottery in Automated Research
Jingjie Ning, Shanshan Zhong, Xiaochuan Li +2
The paper studies how automated research systems can draw misleading conclusions when they rely on a single implementation of an idea, introducing the concept of an "implementation…
Auto Research for Materials: Auditable AI-Scientist Workflows with Held-Out Transfer
Jingjie Ning, Xiaochuan Li, Shanshan Zhong +2
Auto Research uses language-model agents to propose, implement, and evaluate machine-learning changes in a closed loop, but is usually judged by its terminal pipeline. A terminal s…
Neural Chain-of-Thought Search: Searching the Optimal Reasoning Path to Enhance Large Language Models
Guoming Ling, Zhongzhan Huang, Yupei Lin +4
Chain-of-Thought reasoning has significantly enhanced the problem-solving capabilities of Large Language Models. Unfortunately, current models generate reasoning steps sequentially…
AgentWebBench: Benchmarking Multi-Agent Coordination in Agentic Web
Shanshan Zhong, Kate Shen, Chenyan Xiong
Agentic Web is an emerging paradigm where autonomous agents help users use online information. As the paradigm develops, content providers are also deploying agents to manage their…
Agent Skills: A Data-Driven Analysis of Claude Skills for Extending Large Language Model Functionality
George Ling, Shanshan Zhong, Richard Huang
Agent skills extend large language model (LLM) agents with reusable, program-like modules that define triggering conditions, procedural logic, and tool interactions. As these skill…
What Generative Search Engines Like and How to Optimize Web Content Cooperatively
Yujiang Wu, Shanshan Zhong, Yubin Kim +1
By employing large language models (LLMs) to retrieve documents and generate natural language responses, Generative Engines, such as Google AI overview and ChatGPT, provide signifi…