Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
PATH-Bench: Path-Dependent Evaluation of Lifelong Agents
Xidong Yang, Xingyi Zhang, Wenhao Li +7
Lifelong LLM agents increasingly adapt through external learning states that store past interactions as retrievable memories or reusable skills, yet existing benchmarks rarely acco…
cs.AI2025
Reinforced Reasoning for Embodied Planning
Di Wu, Jiaxin Fan, Junzhe Zang +4
Embodied planning requires agents to make coherent multi-step decisions based on dynamic visual observations and natural language goals. While recent vision-language models (VLMs)…