Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
To Nuke or Not to Nuke: LLMs' (Missing) Ethical Reasoning and Actions in a High-Stakes Decision-Making Simulation
John Chen, Sihan Cheng, Can Gurkan +1
Large language models (LLMs) are increasingly deployed as long-horizon agents with decision-making capacities. While LLMs can show ethical competence on dilemmas such as trolley pr…
cs.AI2026
CivBench: Progress-Based Evaluation for LLMs' Strategic Decision-Making in Civilization V
John Chen, Sihan Cheng, Can Gurkan +1
Evaluating strategic decision-making in LLM-based agents requires generative, competitive, and longitudinal environments, yet few benchmarks provide all three, and fewer still offe…
cs.AI2025
Vox Deorum: A Hybrid LLM Architecture for 4X / Grand Strategy Game AI -- Lessons from Civilization V
John Chen, Sihan Cheng, Can Gurkan +2
Large Language Models' capacity to reason in natural language makes them uniquely promising for 4X and grand strategy games, enabling more natural human-AI gameplay interactions su…