Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Deep Networks Learn Deep Hierarchical Models
Amit Daniely
We consider supervised learning with labels and show that layerwise SGD on residual networks can efficiently learn a class of hierarchical models. This model class assumes the…
cs.LG2023
Locally Optimal Descent for Dynamic Stepsize Scheduling
Gilad Yehudai, Alon Cohen, Amit Daniely +3
We introduce a novel dynamic learning-rate scheduling scheme grounded in theory with the goal of simplifying the manual and time-consuming tuning of schedules in practice. Our appr…