16 citations · 24 across the 12 of their papers we have counts for
Showing 2018Show all
2 papers · 1 filter
cs.LG2018
Learning One-hidden-layer Neural Networks under General Input Distributions
Weihao Gao, Ashok Vardhan Makkuva, Sewoong Oh +1
Significant advances have been made recently on training neural networks, where the main challenge is in solving an optimization problem with abundant critical points. However, exi…
cs.LG2018
Breaking the gridlock in Mixture-of-Experts: Consistent and Efficient Algorithms
Ashok Vardhan Makkuva, Sewoong Oh, Sreeram Kannan +1
Mixture-of-Experts (MoE) is a widely popular model for ensemble learning and is a basic building block of highly successful modern neural networks as well as a component in Gated R…