Publications (36)
Word2Minecraft: Generating 3D Game Levels through Large Language Models
Shuo Huang, Muhammad Umair Nasir, Steven James +1
We present Word2Minecraft, a system that leverages large language models to generate playable game levels in Minecraft based on structured stories. The system transforms narrative…
MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents
Simon Rosen, Siddarth Singh, Ebenezer Gelo +6
Evaluating moral alignment in agents navigating conflicting, hierarchically structured human norms is a critical challenge at the intersection of AI safety, moral philosophy, and c…
Investigating Transfer Learning in Graph Neural Networks
Nishai Kooverjee, Steven James, Terence van Zyl
Graph neural networks (GNNs) build on the success of deep learning models by extending them for use in graph spaces. Transfer learning has proven extremely successful for tradition…
Automatic Encoding and Repair of Reactive High-Level Tasks with Learned Abstract Representations
Adam Pacheck, Steven James, George Konidaris +1
We present a framework that, given a set of skills a robot can perform, abstracts sensor data into symbols that we use to automatically encode the robot's capabilities in Linear Te…
Learning Options from Demonstration using Skill Segmentation
Matthew Cockcroft, Shahil Mawjee, Steven James +1
We present a method for learning options from segmented demonstration trajectories. The trajectories are first segmented into skills using nonparametric Bayesian clustering and a r…
CORDA: A Benchmark for Hierarchical Harm-Centric Moral Reasoning in Large Language Models
Siddarth Singh, Victoria Williams, Simon Rosen +6
The key question in moral judgement is not simply whether someone chooses the "right" answer, but how they decide what matters most when moral principles conflict. Current evaluati…