Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent
Dominik Meier, Jan Philip Wahle, Paul Röttger +2
As large language models (LLMs) become integrated into sensitive workflows, concerns grow over their potential to leak confidential information. We propose TrojanStego, a novel thr…
cs.CL2025
Towards Human Understanding of Paraphrase Types in Large Language Models
Dominik Meier, Jan Philip Wahle, Terry Ruas +1
Paraphrases represent a human's intuitive ability to understand expressions presented in various different ways. Current paraphrase evaluations of language models primarily use bin…