5 citations · 8 across the 3 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.CL2024
Risk and Response in Large Language Models: Evaluating Key Threat Categories
Bahareh Harandizadeh, Abel Salinas, Fred Morstatter
This paper explores the pressing issue of risk assessment in Large Language Models (LLMs) as they become increasingly prevalent in various applications. Focusing on how reward mode…
cs.CL2024★ 5 cited
The Butterfly Effect of Altering Prompts: How Small Changes and Jailbreaks Affect Large Language Model Performance
Abel Salinas, Fred Morstatter
Large Language Models (LLMs) are regularly being used to label data across many domains and for myriad tasks. By simply asking the LLM for an answer, or ``prompting,'' practitioner…