1 citations · 1 across the 3 of their papers we have counts for
3 papers
Moral Safety in LLMs: Exposing Performative Compliance with Puzzled Cues
Mohammadamin Shafiei, Shuyue Stella Li, Yulia Tsvetkov
As large language models take on morally consequential roles in healthcare, legal, and hiring contexts, we need to examine whether their ethical behaviors are genuine or superficia…
More or Less Wrong: A Benchmark for Directional Bias in LLM Comparative Reasoning
Mohammadamin Shafiei, Hamidreza Saffari, Nafise Sadat Moosavi
Large language models (LLMs) are known to be sensitive to input phrasing, but the mechanisms by which semantic cues shape reasoning remain poorly understood. We investigate this ph…
MultiHoax: A Dataset of Multi-hop False-Premise Questions
Mohammadamin Shafiei, Hamidreza Saffari, Nafise Sadat Moosavi
As Large Language Models are increasingly deployed in high-stakes domains, their ability to detect false assumptions and reason critically is crucial for ensuring reliable outputs.…