3 papers
cs.AI2026
RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems
Yarin Yerushalmi Levi, Roy Betser, Amit Giloni +5
Agentic AI systems powered by large language models (LLMs) are rapidly evolving into autonomous decision-making systems, exposing attack vectors beyond those of traditional LLM vul…
cs.CV2026
Identifying Memorization of Diffusion Models through -Laplace Analysis: Estimators, Bounds and Applications
Jonathan Brokman, Itay Gershon, Amit Giloni +4
Diffusion models, today's leading image generative models, estimate the score function, i.e. the gradient of the log probability of (perturbed) data samples, without direct access…
cs.LG2026
Training-Free Policy Violation Detection via Activation-Space Whitening in LLMs
Oren Rachmil, Avishag Shapira, Roy Betser +5
As organizations increasingly deploy LLMs in sensitive domains such as legal, financial, and medical settings, ensuring alignment with internal organizational policies has become a…