2 papers
cs.RO2026
SHRIMP: Iterative Refinement of Robot Task Plans
Mya Schroder, Yuna Hwang, Callie Y. Kim +5
As collaborative robots have entered domains such as manufacturing, agriculture, and healthcare, programming or adapting robot behavior typically requires robotic expertise that mo…
cs.AI2026
The Long-Horizon Task Mirage? Diagnosing Where and Why Agentic Systems Break
Xinyu Jessica Wang, Haoyue Bai, Yiyou Sun +7
Large language model (LLM) agents perform strongly on short- and mid-horizon tasks, but often break down on long-horizon tasks that require extended, interdependent action sequence…