Contributions to Large Scale Bayesian Inference and Adversarial Machine Learning
arXiv:2109.13232
Abstract
The rampant adoption of ML methodologies has revealed that models are usually adopted to make decisions without taking into account the uncertainties in their predictions. More critically, they can be vulnerable to adversarial examples. Thus, we believe that developing ML systems that take into account predictive uncertainties and are robust against adversarial examples is a must for critical, real-world tasks. We start with a case study in retailing. We propose a robust implementation of the Nerlove-Arrow model using a Bayesian structural time series model. Its Bayesian nature facilitates incorporating prior information reflecting the manager's views, which can be updated with relevant data. However, this case adopted classical Bayesian techniques, such as the Gibbs sampler. Nowadays, the ML landscape is pervaded with neural networks and this chapter also surveys current developments in this sub-field. Then, we tackle the problem of scaling Bayesian inference to complex models and large data regimes. In the first part, we propose a unifying view of two different Bayesian inference algorithms, Stochastic Gradient Markov Chain Monte Carlo (SG-MCMC) and Stein Variational Gradient Descent (SVGD), leading to improved and efficient novel sampling schemes. In the second part, we develop a framework to boost the efficiency of Bayesian inference in probabilistic models by embedding a Markov chain sampler within a variational posterior approximation. After that, we present an alternative perspective on adversarial classification based on adversarial risk analysis, and leveraging the scalable Bayesian approaches from chapter 2. In chapter 4 we turn to reinforcement learning, introducing Threatened Markov Decision Processes, showing the benefits of accounting for adversaries in RL while the agent learns.
PhD thesis
References in corpus (23)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- PyTorch: An Imperative Style, High-Performance Deep Learning Library
- Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling
- Fashion-MNIST: a Novel Image Dataset for Benchmarking Machine Learning Algorithms
- ADADELTA: An Adaptive Learning Rate Method
- Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
- Language Models are Few-Shot Learners
- Semi-Supervised Learning with Deep Generative Models
- Inferring causal impact using Bayesian structural time-series models
- Security Evaluation of Pattern Classifiers under Attack
- Robust Adversarial Reinforcement Learning
- Markov Chain Monte Carlo and Variational Inference: Bridging the Gap
- A General Framework for the Parametrization of Hierarchical Models
- Deep Learning: A Bayesian Perspective
- AI Safety Gridworlds
- Dropout Inference in Bayesian Neural Networks with Alpha-divergences
- Mathematics of Deep Learning
- Variational Inference using Implicit Distributions
- NeuTra-lizing Bad Geometry in Hamiltonian Monte Carlo Using Neural Transport
- A Selective Overview of Deep Learning
- Learning to Draw Samples with Amortized Stein Variational Gradient Descent
- A Contrastive Divergence for Combining Variational Inference and MCMC
- A Statistician Teaches Deep Learning