most citedMirrorMind: Empowering OmniScientist with the Expert Perspectives and Collective Knowledge of Human Scientists

1 citations · 1 across the 5 of their papers we have counts for

collaborators

7 papers

cs.AI2026

Ideation Arena: Evaluating LLM Generated Research Ideas with Battle-style Human Expert Assessment

Zhiyu Chen, Keyu Zhao, Jigao Fu +8

Evaluating research ideas generated by LLMs is difficult because their scientific value cannot be fully determined by objective criteria, and no single reference answer specifies w…

cs.AI2026

LitReview Arena: Evaluating Literature Review Agents with Battle-Style Peer Review Platform

Ruotong Zhao, Zhiyu Chen, Xurui Liu +7

Literature reviews are essential to scientific progress, but rigorously evaluating automatically generated reviews remains difficult because many aspects of research utility depend…

cs.AI2026

Clarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific Collaboration

Zihan Guo, Zeyi Chen, Zhiyu Chen +15

Existing autonomous research agents can support parts of the research process, but most systems still treat research as either an isolated assistant task or a closed workflow. Ther…

cs.AI2026

SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior

Zhiyu Chen, Zihan Guo, Bo Huang +4

Agent Skills augment large language model (LLM) agents with procedural knowledge at inference time, but current benchmarks rarely distinguish what a Skill says from how it is organ…

cs.CR2026

SkillProbe: Security Auditing for Emerging Agent Skill Marketplaces via Multi-Agent Collaboration

Zihan Guo, Zhiyu Chen, Xiaohang Nie +3

With the rapid evolution of Large Language Model (LLM) agent ecosystems, centralized skill marketplaces have emerged as pivotal infrastructure for augmenting agent capabilities. Ho…

cs.CY2025

OmniScientist: Toward a Co-evolving Ecosystem of Human and AI Scientists

Chenyang Shao, Dehao Huang, Yu Li +18

With the rapid development of Large Language Models (LLMs), AI agents have demonstrated increasing proficiency in scientific tasks, ranging from hypothesis generation and experimen…