3 papers
cs.AI2026
Supervised sparse auto-encoders for interpretable and compositional representations
Ouns El Harzli, Hugo Wallner, Yoonsoo Nam +1
Sparse auto-encoders (SAEs) have re-emerged as a prominent method for mechanistic interpretability, yet they face two significant challenges: the non-smoothness of the penalt…
cs.LG2026
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
Ouns El Harzli, Yoonsoo Nam, Ilja Kuzborskij +2
Algorithmic stability is a classical framework for analyzing the generalization error of learning algorithms. It predicts that an algorithm has small generalization error if it is…
cs.AI2025
From Neural Networks to Logical Theories: The Correspondence between Fibring Modal Logics and Fibring Neural Networks
Ouns El Harzli, Bernardo Cuenca Grau, Artur d'Avila Garcez +2
Fibring of modal logics is a well-established formalism for combining countable families of modal logics into a single fibred language with common semantics, characterized by fibre…