From the 1 of 3 linked papers with an AI index.
3 papers
ToxScreen: Detecting Whether an LLM Has Been Poisoned
Anthony Hughes, Nicole Xing, Collin Francel +2
The paper introduces ToxScreen, a benchmark of backdoored large language models, and evaluates methods for recovering hidden triggers under realistic defender constraints, finding…
Phantom Transfer: Data Poisoning can Survive Data-Level Defences
Andrew Draganov, Tolga H. Dur, Anandmayi Bhongade +1
We present a data poisoning attack -- Phantom Transfer -- with the property that, even if you know precisely how the poison was placed into an otherwise benign dataset, you cannot…
Ultrametric Cluster Hierarchies: I Want 'em All!
Andrew Draganov, Pascal Weber, Rasmus Skibdahl Melanchton Jørgensen +3
Hierarchical clustering is a powerful tool for exploratory data analysis, organizing data into a tree of clusterings from which a partition can be chosen. This paper generalizes th…