1 citations · 3 across the 5 of their papers we have counts for
4 papers · 1 filter
The Topos of Transformer Networks
Mattia Jacopo Villani, Peter McBurney
The transformer neural network has significantly out-shined all other neural network architectures as the engine behind large language models. We provide a theoretical analysis of…
Learning Translations: Emergent Communication Pretraining for Cooperative Language Acquisition
Dylan Cope, Peter McBurney
In Emergent Communication (EC) agents learn to communicate with one another, but the protocols that they develop are specialised to their training community. This observation led t…
Joining the Conversation: Towards Language Acquisition for Ad Hoc Team Play
Dylan Cope, Peter McBurney
In this paper, we propose and consider the problem of cooperative language acquisition as a particular form of the ad hoc team play problem. We then present a probabilistic model f…
Unwrapping All ReLU Networks
Mattia Jacopo Villani, Peter McBurney
Deep ReLU Networks can be decomposed into a collection of linear models, each defined in a region of a partition of the input space. This paper provides three results extending thi…