2 papers
cs.AI2026
Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Remember
Zenghuang Fu, Zhaoyang Li, Qiuyuan Ai +6
Self-play agents can generate training problems without questions from target benchmarks, but their curricula lack persistent state: failures affect gradients yet do not explicitly…
cs.AI2026
UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following
Kun Yu, Jianhua Yang, Yixiang Chen +7
Language-guided human following is an important capability for embodied agents, but existing benchmarks typically assume that the target person is visible at the start of an episod…