229 citations · 547 across the 29 of their papers we have counts for
Showing 2023Show all
2 papers · 1 filter
cs.LG2023
The Feature Speed Formula: a flexible approach to scale hyper-parameters of deep neural networks
Lénaïc Chizat, Praneeth Netrapalli
Deep learning succeeds by doing hierarchical feature learning, yet tuning hyper-parameters (HP) such as initialization scales, learning rates etc., only give indirect control over…
stat.ML2023
Near Optimal Heteroscedastic Regression with Symbiotic Learning
Dheeraj Baby, Aniket Das, Dheeraj Nagaraj +1
We consider the problem of heteroscedastic linear regression, where, given samples from $y_i = \langle \mathbf{w}^{*}, \mathbf{x}_i \rangle + ε_i \cdot \l…