3 papers
cs.CR2026
Attacks and Mitigations for Distributed Governance of Agentic AI under Byzantine Adversaries
Matthew D. Laws, Alina Oprea, Cristina Nita-Rotaru
Agentic AI governance is a critical component of agentic AI infrastructure ensuring that agents follow their owner's communication and interaction policies, and providing protectio…
cs.CR2026
Reconstruction of Personally Identifiable Information from Proprietary Data in Supervised Fine-Tuned Models
Sae Furukawa, Alina Oprea
Supervised Finetuning (SFT) has become one of the primary methods for adapting a large language model (LLM) with extensive pre-trained knowledge to domain-specific, instruction-fol…
cs.CR2026
Toward a Principled Framework for Agent Safety Measurement
Shuyi Lin, Anshuman Suri, Alina Oprea +1
LLM agents emit actions, not just text, and once taken, those actions often cannot be undone. Yet today's agent-safety evaluations run greedy or a few sampled rollouts and report a…