8 citations · 15 across the 9 of their papers we have counts for
Showing 2022Show all
2 papers · 1 filter
cs.LG2022
Generalisation and the Risk--Entropy Curve
Dominic Belcher, Antonia Marcu, Adam Prügel-Bennett
In this paper we show that the expected generalisation performance of a learning machine is determined by the distribution of risks or equivalently its logarithm -- a quantity we t…
cs.LG2022
Orthogonalising gradients to speed up neural network optimisation
Mark Tuddenham, Adam Prügel-Bennett, Jonathan Hare
The optimisation of neural networks can be sped up by orthogonalising the gradients before the optimisation step, ensuring the diversification of the learned representations. We or…