1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2025
Beyond Mimicry: Preference Coherence in LLMs
Luhan Mikaelson, Derek Shiller, Hayley Clatterbuck
We investigate whether large language models exhibit genuine preference structures by testing their responses to AI-specific trade-offs involving GPU reduction, capability restrict…
cs.LG2025★ 1 cited
Self-Ablating Transformers: More Interpretability, Less Sparsity
Jeremias Ferrao, Luhan Mikaelson, Keenan Pepper +1
A growing intuition in machine learning suggests a link between sparsity and interpretability. We introduce a novel self-ablation mechanism to investigate this connection ante-hoc…