autonomous multi-agent systems 1black-box extraction 1defense mechanisms 1inference-time harness 1ip leakage 1llm security 1
From the 1 of 12 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents
Tianjun Pan, Yuan Li, Hongda Wang +8
External natural-language skills provide large language model (LLM) agents with reusable and editable guidance for solving complex tasks. Yet their effectiveness depends not only o…
cs.AI2026
Segment-Aligned Policy Optimization for Multi-Modal Reasoning
Lei Gao, Zhuoming Li, Mengxi Jia +4
Existing reinforcement learning approaches for Large Language Models typically perform policy optimization at the granularity of individual tokens or entire response sequences. How…