Showing cs.CYShow all
2 papers · 1 filter
cs.CY2026
Simple Role Assignment is Extraordinarily Effective for Safety Alignment
Zhou Ziheng, Jiakun Ding, Zhaowei Zhang +6
Principle-based alignment often lacks context sensitivity and completeness. Grounded in Theory of Mind, we propose role conditioning as a compact alternative: social roles (e.g., m…
cs.CY2025
A Human-centric Framework for Debating the Ethics of AI Consciousness Under Uncertainty
Zhou Ziheng, Haiqiang Dai, Bin Ling +2
As AI systems become increasingly sophisticated, questions about machine consciousness and its ethical implications have moved from fringe speculation to mainstream academic debate…