Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Hidden in the Request: Explaining Unethical LLM Compliance through Token Relevance
Or Biton, Tomer Krichli, Itai Allouche +1
Although Large Language Models (LLMs) are aligned to optimize for both helpfulness and harmlessness, these dual objectives may conflict, inevitably leading to alignment failures. T…
cs.AI2026
Why Sampling Is Not Choosing: Intentionality, Agency, and Moral Responsibility in Large Language Models
Joseph Keshet
Recent advances in large language models (LLMs) have prompted claims that such systems exhibit agency or qualify as moral agents. This paper argues that these attributions are misg…