Showing stat.MLShow all
2 papers · 1 filter
stat.ML2024
Fast training of large kernel models with delayed projections
Amirhesam Abedsoltan, Siyuan Ma, Parthe Pandit +1
Classical kernel machines have historically faced significant challenges in scaling to large datasets and model sizes--a key ingredient that has driven the success of neural networ…
stat.ML2018
Kernel machines that adapt to GPUs for effective large batch training
Siyuan Ma, Mikhail Belkin
Modern machine learning models are typically trained using Stochastic Gradient Descent (SGD) on massively parallel computing resources such as GPUs. Increasing mini-batch size is a…