collaborators

7 papers

cs.AI2026

Prototype Transformer: Towards Language Model Architectures Interpretable by Design

Yordan Yordanov, Matteo Forasassi, Bayar Menzat +6

While state-of-the-art language models (LMs) surpass most humans in certain domains, their reasoning remains largely opaque, reducing trust and increasing the risk of deception and…

cs.LG2026

Faster Predictive Coding Networks via Better Initialization

Luca Pinchetti, Simon Frieder, Thomas Lukasiewicz +1

Research aimed at scaling up neuroscience inspired learning algorithms for neural networks is accelerating. Recently, a key research area has been the study of energy-based learnin…

cs.LG2025

Data for Mathematical Copilots: Better Ways of Presenting Proofs for Machine Learning

Simon Frieder, Jonas Bayer, Sam Looi +13

The datasets and benchmarks commonly used to train and evaluate the mathematical capabilities of AI-based mathematical copilots (primarily large language models) exhibit several sh…

cs.LG2025

Towards the Training of Deeper Predictive Coding Neural Networks

Chang Qi, Matteo Forasassi, Thomas Lukasiewicz +1

Predictive coding networks are neural models that perform inference through an iterative energy minimization process, whose operations are local in space and time. While effective…

cs.CL2025

Shh, don't say that! Domain Certification in LLMs

Cornelius Emde, Alasdair Paren, Preetham Arvind +6

Large language models (LLMs) are often deployed to perform constrained tasks, with narrow domains. For example, customer support bots can be built on top of LLMs, relying on their…

cs.LG2025

Towards Certification of Uncertainty Calibration under Adversarial Attacks

Cornelius Emde, Francesco Pinto, Thomas Lukasiewicz +2

Since neural classifiers are known to be sensitive to adversarial perturbations that alter their accuracy, \textit{certification methods} have been developed to provide provable gu…