1 paper
Francois Caron, Fadhel Ayed, Paul Jung +3
We consider gradient-based optimisation of wide, shallow neural networks, where the output of each hidden node is scaled by a positive parameter. The scaling parameters are non-ide…