Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Frame-Conditioned Moral Computation in LLaMA 3.1-8B-Instruct: A Mechanistic Interpretability Audit of Ethical Reasoning
Ali Dasdan, Manan Shah, W. Russell Neuman +3
Behavioral audits of Large Language Models on moral prompts measure what the model says, not the internal computation producing it. We use Transluce, an AI-driven mechanistic-inter…
cs.AI2026
Six Llamas: Comparative Religious Ethics Through LoRA-Adapted Language Models
Chad Coleman, W. Russell Neuman, Manan Shah +5
We present Six Llamas, a comparative study examining whether large language models fine-tuned on distinct religious corpora encode systematically different patterns of ethical reas…