3 papers
cs.CL2026
Unsupervised Elicitation of Moral Values from Language Models
Meysam Alizadeh, Fabrizio Gilardi, Zeynab Samei
As AI systems become pervasive, grounding their behavior in human values is critical. Prior work suggests that language models (LMs) exhibit limited inherent moral reasoning, leadi…
cs.CL2025
Web-Browsing LLMs Can Access Social Media Profiles and Infer User Demographics
Meysam Alizadeh, Fabrizio Gilardi, Zeynab Samei +1
Large language models (LLMs) have traditionally relied on static training data, limiting their knowledge to fixed snapshots. Recent advancements, however, have equipped LLMs with w…
cs.CR2025
Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution
Meysam Alizadeh, Zeynab Samei, Daria Stetsenko +1
Previous benchmarks on prompt injection in large language models (LLMs) have primarily focused on generic tasks and attacks, offering limited insights into more complex threats lik…