10 papers
EvalMORAAL: Interpretable Chain-of-Thought and LLM-as-Judge Evaluation for Moral Alignment in Large Language Models
Hadi Mohammadi, Anastasia Giachanou, Robert A. Bagheri
We present EvalMORAAL, a transparent chain-of-thought (CoT) framework that uses two scoring methods (log-probabilities and direct ratings) plus a model-as-judge peer review to eval…
Correspondence Analysis and PMI-Based Word Embeddings: A Comparative Study
Qianqian Qi, Ayoub Bagheri, David J. Hessen +1
Popular word embedding methods such as GloVe and Word2Vec are related to the factorization of the pointwise mutual information (PMI) matrix. In this paper, we establish a formal co…
Mind the Style: Impact of Communication Style on Human-Chatbot Interaction
Erik Derner, Dalibor KuÄera, Dalibor Kučera +3
Conversational agents increasingly mediate everyday digital interactions, yet the effects of their communication style on user experience and task success remain insufficiently und…
Exploring Cultural Variations in Moral Judgments with Large Language Models
Hadi Mohammadi, Ayoub Bagheri
Large Language Models (LLMs) have shown strong performance across many tasks, but their ability to capture culturally diverse moral values remains unclear. In this paper, we examin…
Explainability-Based Token Replacement on LLM-Generated Text
Hadi Mohammadi, Anastasia Giachanou, Daniel L. Oberski +1
Generative models, especially large language models (LLMs), have shown remarkable progress in producing text that appears human-like. However, they often exhibit patterns that make…
Do Large Language Models Understand Morality Across Cultures?
Hadi Mohammadi, Yasmeen F. S. S. Meijer, Efthymia Papadopoulou +1
Recent advancements in large language models (LLMs) have established them as powerful tools across numerous domains. However, persistent concerns about embedded biases, such as gen…