papers
Publications (9)
cs.LG2024
Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Avi Singh, John D. Co-Reyes, Rishabh Agarwal +38
cs.LG2019
Dueling Decoders: Regularizing Variational Autoencoder Latent Spaces
Bryan Seybold, Emily Fertig, Alex Alemi +1
cs.CL2023
Frontier Language Models are not Robust to Adversarial Arithmetic, or "What do I need to say so you agree 2+2=5?
C. Daniel Freeman, Laura Culp, Aaron Parisi +27
cs.LG2017
TensorFlow Distributions
Joshua V. Dillon, Ian Langmore, Dustin Tran +7
cs.CL2024
Training LLMs over Neurally Compressed Text
Brian Lester, Jaehoon Lee, Alex Alemi +4
cs.LG2018
Watch Your Step: Learning Node Embeddings via Graph Attention
Sami Abu-El-Haija, Bryan Perozzi, Rami Al-Rfou +1
cs.CV2017
Motion Prediction Under Multimodality with Conditional Stochastic Networks
Katerina Fragkiadaki, Jonathan Huang, Alex Alemi +3
cs.CV2016
Inception-v4, Inception-ResNet and the Impact of Residual Connections on Learning
Christian Szegedy, Sergey Ioffe, Vincent Vanhoucke +1
cs.LG2023
Small-scale proxies for large-scale Transformer training instabilities
Mitchell Wortsman, Peter J. Liu, Lechao Xiao +13