18 citations · 30 across the 3 of their papers we have counts for
3 papers
cs.CL2023★ 18 cited
Exploring the MIT Mathematics and EECS Curriculum Using Large Language Models
Sarah J. Zhang, Samuel Florin, Ariel N. Lee +12
We curate a comprehensive dataset of 4,550 questions and solutions from problem sets, midterm exams, and final exams across all MIT Mathematics and Electrical Engineering and Compu…
stat.ML2022
Plateau in Monotonic Linear Interpolation -- A "Biased" View of Loss Landscape for Deep Networks
Xiang Wang, Annie N. Wang, Mo Zhou +1
Monotonic linear interpolation (MLI) - on the line connecting a random initialization with the minimizer it converges to, the loss and accuracy are monotonic - is a phenomenon that…
cs.LG2020★ 12 cited
Dissecting Hessian: Understanding Common Structure of Hessian in Neural Networks
Yikai Wu, Xingyu Zhu, Chenwei Wu +2
Hessian captures important properties of the deep neural network loss landscape. Previous works have observed low rank structure in the Hessians of neural networks. In this paper,…