From the 1 of 7 linked papers with an AI index.
7 papers
EMAS: Stabilizing Multi-Agent System Evolution through Evidence-Guided Revision
Chao Fei, Qingyi Si, Kaihua Liang +3
Many methods for automated multi-agent system design optimize prompts and topologies during an initial design stage and then deploy the resulting system unchanged on subsequent sam…
A Self-Evolving Agent for Longitudinal Personal Health Management
Haoran Li, Jiebi Deng, Tong Jin +10
The paper presents HealthClaw, an open‑source AI agent that maintains and updates a personal health profile over time, improving answer accuracy and privacy compared to standard pr…
FORGE: Research-Trajectory Hijacking Attacks on Deep Research Agents
Yue Pan, Ziheng Zhang, Junxiang Lei +3
Deep research agents decompose open-ended queries into subtasks, retrieve web evidence over multiple rounds, and synthesize long-form reports. This workflow creates a planning-laye…
Evaluating Stochastic Collapse and Implicit Bias in Multimodal Large Language Models
Huiyuan Zheng, Houtao Zhang, Boyang Wang +2
Current evaluations for Multimodal Large Language Models (MLLMs) overwhelmingly focus on utility-driven objectives, leaving model behavior under logic-neutral scenarios largely und…
Outcome-Grounded Advantage Reshaping for Fine-Grained Credit Assignment in Mathematical Reasoning
Ziheng Li, Liu Kang, Feng Xiao +7
Group Relative Policy Optimization (GRPO) has emerged as a promising critic-free reinforcement learning paradigm for reasoning tasks. However, standard GRPO employs a coarse-graine…
GUI Agents for Continual Game Generation
Yixu Huang, Bo Li, Na Li +8
Generating a game is not the same as making one that can be played. Despite advances in code generation, existing approaches treat game generation as one-shot translation from prom…