collaborators

16 papers

cs.AI2026

AREX: Towards a Recursively Self-Improving Agent for Deep Research

Shuqi Lu, Chaofan Li, Kun Luo +21

Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candidate can often be decomposed i…

cs.AI2026

IntElicit: Eliciting and Assessing Contextualized Creativity via Dialogue Policy Optimization

Mingjia Li, Jin Wu, Hong Qian +7

Contextualized assessment offers high ecological validity for evaluating creativity but introduces a critical challenge: observed performance may be confounded with cognitive profi…

cs.CL2026

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

Jiajie Jin, Yuyang Hu, Kai Qiu +15

Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the evidence, and carry the result…

cs.AI2026

PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Training Large Language Models

Ziliang Zhao, Zenan Xu, Shuting Wang +7

Planning is a fundamental capability for large language models (LLMs) because such complex tasks require models to coordinate goals, constraints, resources, and long-term consequen…

cs.AI2026

AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning

Yuyang Hu, Hongjin Qian, Shuting Wang +5

Recent progress on long-horizon agentic tasks has been driven largely by scaling up individual agents through stronger models, better tools, and more effective scaffolding. In cont…

cs.AI2026

SAM: State-Adaptive Memory for Long-Horizon Reasoning Agent

Yuyang Hu, Hongjin Qian, Shuting Wang +5

Long-horizon agentic reasoning requires large language models to act over long interaction histories containing thoughts, tool calls, observations, and partial conclusions. The cha…