4 citations · 5 across the 12 of their papers we have counts for
3 papers · 1 filter
A Geometric Unification of Concept Learning with Concept Cones
Alexandre Rocchi, Thomas Fel, Gianni Franchi
Two traditions of interpretability have evolved side by side but seldom spoken to each other: Concept Bottleneck Models (CBMs), which prescribe what a concept should be, and Sparse…
The Case for Model Science: Verify, Explore, Steer, Refine
Przemyslaw Biecek, Luca Longo, Jianlong Zhou +3
We argue that the AI community is now ready to move beyond benchmarking and consolidate scattered efforts in model analysis into a systematic discipline, a direction we term Model…
Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts
Sheridan Feucht, Tal Haklay, Usha Bhalla +9
Does structure in representations imply structure in computation? We study how Llama-3.1-8B reasons over cyclic concepts (e.g., "what month is six months after August?"). Even thou…