4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CR2023★ 4 cited
Ignore This Title and HackAPrompt: Exposing Systemic Vulnerabilities of LLMs through a Global Scale Prompt Hacking Competition
Sander Schulhoff, Jeremy Pinto, Anaum Khan +7
Large Language Models (LLMs) are deployed in interactive contexts with direct user engagement, such as chatbots and writing assistants. These deployments are vulnerable to prompt i…
cs.CV2023
Coloring Deep CNN Layers with Activation Hue Loss
Louis-François Bouchard, Mohsen Ben Lazreg, Matthew Toews
This paper proposes a novel hue-like angular parameter to model the structure of deep convolutional neural network (CNN) activation space, referred to as the {\em activation hue},…