3 papers
cs.LG2026
Weaves, Wires, and Morphisms: Formalizing and Implementing the Algebra of Deep Learning
Vincent Abbott, Gioele Zardini
Despite deep learning models running well-defined mathematical functions, we lack a formal mathematical framework for describing model architectures. Ad-hoc notation, diagrams, and…
math.CT2025
Accelerating Machine Learning Systems via Category Theory: Applications to Spherical Attention for Gene Regulatory Networks
Vincent Abbott, Kotaro Kamiya, Gerard Glowacki +3
How do we enable artificial intelligence models to improve themselves? This is central to exponentially improving generalized artificial intelligence models, which can improve thei…
cs.LG2025
FlashAttention on a Napkin: A Diagrammatic Approach to Deep Learning IO-Awareness
Vincent Abbott, Gioele Zardini
Optimizing deep learning algorithms currently requires slow, manual derivation, potentially leaving much performance untapped. Methods like FlashAttention have achieved a x6 perfor…