30 citations · 38 across the 3 of their papers we have counts for
3 papers
LLM Defenses Are Not Robust to Multi-Turn Human Jailbreaks Yet
Nathaniel Li, Ziwen Han, Ian Steneker +6
Recent large language model (LLM) defenses have greatly improved models' ability to refuse harmful queries, even when adversarially attacked. However, LLM defenses are primarily ev…
The Drift of #MyBodyMyChoice Discourse on Twitter
Cristina Menghini, Justin Uhr, Shahrzad Haddadan +3
#MyBodyMyChoice is a well-known hashtag originally created to advocate for women's rights, often used in discourse about abortion and bodily autonomy. The Covid-19 outbreak prompte…
RePBubLik: Reducing the Polarized Bubble Radius with Link Insertions
Shahrzad Haddadan, Cristina Menghini, Matteo Riondato +1
The topology of the hyperlink graph among pages expressing different opinions may influence the exposure of readers to diverse content. Structural bias may trap a reader in a polar…