18 citations · 18 across the 8 of their papers we have counts for
4 papers · 1 filter
Adversarial Reasoning at Jailbreaking Time
Mahdi Sabbaghi, Paul Kassianik, George Pappas +3
As large language models (LLMs) are becoming more capable and widespread, the study of their failure cases is becoming increasingly important. Recent advances in standardizing, mea…
Extracting Memorized Training Data via Decomposition
Ellen Su, Anu Vellore, Amy Chang +4
The widespread use of Large Language Models (LLMs) in society creates new information security challenges for developers, organizations, and end-users alike. LLMs are trained on la…
Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
Anay Mehrotra, Manolis Zampetakis, Paul Kassianik +4
While Large Language Models (LLMs) display versatile functionality, they continue to generate harmful, biased, and toxic content, as demonstrated by the prevalence of human-designe…
Merlion: A Machine Learning Library for Time Series
Aadyot Bhatnagar, Paul Kassianik, Chenghao Liu +20
We introduce Merlion, an open-source machine learning library for time series. It features a unified interface for many commonly used models and datasets for anomaly detection and…