Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
A quantitative analysis of semantic information in deep representations of text and images
Santiago Acevedo, Andrea Mascaretti, Riccardo Rende +3
It was recently observed that the representations of different models that process identical or semantically related inputs tend to align. We analyze this phenomenon using the Info…
cs.CL2025
A distributional simplicity bias in the learning dynamics of transformers
Riccardo Rende, Federica Gerace, Alessandro Laio +1
The remarkable capability of over-parameterised neural networks to generalise effectively has been explained by invoking a ``simplicity bias'': neural networks prevent overfitting…