238 citations · 356 across the 7 of their papers we have counts for
7 papers
Diffusion-LM Improves Controllable Text Generation
Xiang Lisa Li, John Thickstun, Ishaan Gulrajani +2
Controlling the behavior of language models (LMs) without re-training is a major open problem in natural language generation. While recent works have demonstrated successes on cont…
Convexified Convolutional Neural Networks
Yuchen Zhang, Percy Liang, Martin J. Wainwright
We describe the class of convexified convolutional neural networks (CCNNs), which capture the parameter sharing of convolutional neural networks in a convex manner. By representing…
Estimation from Indirect Supervision with Linear Moments
Aditi Raghunathan, Roy Frostig, John Duchi +1
In structured prediction problems where we have indirect supervision of the output, maximum marginal likelihood faces two computational obstacles: non-convexity of the objective an…
Tensor Factorization via Matrix Factorization
Volodymyr Kuleshov, Arun Tejasvi Chaganty, Percy Liang
Tensor factorization arises in many machine learning applications, such knowledge base modeling and parameter estimation in latent variable models. However, numerical methods for t…
The Statistics of Streaming Sparse Regression
Jacob Steinhardt, Stefan Wager, Percy Liang
We present a sparse analogue to stochastic gradient descent that is guaranteed to perform well under similar conditions to the lasso. In the linear regression setup with irrepresen…
Altitude Training: Strong Bounds for Single-Layer Dropout
Stefan Wager, William Fithian, Sida Wang +1
Dropout training, originally designed for deep neural networks, has been successful on high-dimensional single-layer natural language tasks. This paper proposes a theoretical expla…