18 citations · 19 across the 10 of their papers we have counts for
Showing 2026 · cs.CLShow all
2 papers · 2 filters
cs.CL2026
MechaTerp-TRACE: A Novel Approach for Component Ablation Analysis in Language Models
Brandon Colelough, Davis Bartels, Madeline Bittner +1
Interpretability research on large language models has produced accounts of factual recall in feed-forward layers and of token relationships in self-attention, but little work offe…
cs.CL2026
Quantifying Hallucinations in Language Language Models on Medical Textbooks
Brandon C. Colelough, Davis Bartels, Dina Demner-Fushman
Hallucinations, the tendency for large language models to provide responses with factually incorrect and unsupported claims, is a serious problem within natural language processing…