2 papers
cs.LG2025
ENIGMA: The Geometry of Reasoning and Alignment in Large-Language Models
Gareth Seneque, Lap-Hang Ho, Nafise Erfanian Saeedi +3
We present Entropic Mutual-Information Geometry Large-Language Model Alignment (ENIGMA), a novel approach to Large-Language Model (LLM) training that jointly improves reasoning, al…
cs.LG2024
ABC Align: Large Language Model Alignment for Safety & Accuracy
Gareth Seneque, Lap-Hang Ho, Ariel Kuperman +2
Alignment of Large Language Models (LLMs) remains an unsolved problem. Human preferences are highly distributed and can be captured at multiple levels of abstraction, from the indi…