1 paper
Diego Dorn, Alexandre Variengien, Charbel-Raphaël Segerie +1
Input-output safeguards are used to detect anomalies in the traces produced by Large Language Models (LLMs) systems. These detectors are at the core of diverse safety-critical appl…