3 papers
cs.AI2026
Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents
Xi Zhang, Meijun Gao, Yuntian Zhao +6
Large Language Model (LLM) agents increasingly act inside real workspaces, where tools and skills determine whether model reasoning becomes reliable action. Existing skills remain…
cs.RO2026
AgentRob: From Virtual Forum Agents to Hijacked Physical Robots
Wenrui Liu, Yaxuan Wang, Xun Zhang +13
Large Language Model (LLM)-powered autonomous agents have demonstrated significant capabilities in virtual environments, yet their integration with the physical world remains narro…
cs.CL2026
How Few-shot Demonstrations Affect Prompt-based Defenses Against LLM Jailbreak Attacks
Yanshu Wang, Shuaishuai Yang, Jingjing He +1
Large Language Models (LLMs) face increasing threats from jailbreak attacks that bypass safety alignment. While prompt-based defenses such as Role-Oriented Prompts (RoP) and Task-O…