2 papers
cs.LG2026
From Perception to Action: Spatial AI Agents and World Models
Gloria Felicia, Nolan Bryant, Handi Putra +3
While large language models have become the prevailing approach for agentic reasoning and planning, their success in symbolic domains does not readily translate to the physical wor…
cs.LG2026
StepShield: When, Not Whether to Intervene on Rogue Agents
Gloria Felicia, Zitha Sasindran, Jinfeng He +3
Agent safety benchmarks measure whether a monitor detects harm, not when. Yet timing is the difference between intervention and autopsy. We introduce StepShield, the first benchmar…