4 papers · 1 filter
Position: Zeroth-Order Optimization in Deep Learning Is Underexplored, Not Underpowered
Sijia Liu, Yicheng Lang, Soumyadeep Pal +6
Zeroth-order (ZO) optimization, learning from finite differences of function evaluations without backpropagation, has recently regained attention in deep learning due to its memory…
BOOM: Benchmarking Out-Of-distribution Molecular Property Predictions of Machine Learning Models
Evan R. Antoniuk, Shehtab Zaman, Tal Ben-Nun +9
Data-driven molecular discovery leverages artificial intelligence/machine learning (AI/ML) and generative modeling to filter and design novel molecules. Discovering novel molecules…
Forecasting Fails: Unveiling Evasion Attacks in Weather Prediction Models
Huzaifa Arif, Pin-Yu Chen, Alex Gittens +2
With the increasing reliance on AI models for weather forecasting, it is imperative to evaluate their vulnerability to adversarial perturbations. This work introduces Weather Adapt…
Trajectory Balance with Asynchrony: Decoupling Exploration and Learning for Fast, Scalable LLM Post-Training
Brian Bartoldson, Siddarth Venkatraman, James Diffenderfer +7
Reinforcement learning (RL) is a critical component of large language model (LLM) post-training. However, on-policy algorithms used for post-training are not naturally robust to a…