10 citations · 10 across the 3 of their papers we have counts for
3 papers
StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
Jillian Fisher, Skyler Hallinan, Ximing Lu +3
Authorship obfuscation, rewriting a text to intentionally obscure the identity of the author, is an important but challenging task. Current methods using large language models (LLM…
Localizing Paragraph Memorization in Language Models
Niklas Stoehr, Mitchell Gordon, Chiyuan Zhang +1
Can we localize the weights and mechanisms used by a language model to memorize and recite entire paragraphs of its training data? In this paper, we show that while memorization is…
Cura: Curation at Social Media Scale
Wanrong He, Mitchell L. Gordon, Lindsay Popowski +1
How can online communities execute a focused vision for their space? Curation offers one approach, where community leaders manually select content to share with the community. Cura…