2 papers
cs.CL2025
Concept Bottleneck Large Language Models
Chung-En Sun, Tuomas Oikarinen, Berk Ustun +1
We introduce Concept Bottleneck Large Language Models (CB-LLMs), a novel framework for building inherently interpretable Large Language Models (LLMs). In contrast to traditional bl…
cs.LG2025
Understanding Fixed Predictions via Confined Regions
Connor Lawless, Tsui-Wei Weng, Berk Ustun +1
Machine learning models can assign fixed predictions that preclude individuals from changing their outcome. Existing approaches to audit fixed predictions do so on a pointwise basi…