Perfect density models cannot guarantee anomaly detection
arXiv:2012.03808 · doi:10.3390/e23121690
Abstract
Thanks to the tractability of their likelihood, several deep generative models show promise for seemingly straightforward but important applications like anomaly detection, uncertainty estimation, and active learning. However, the likelihood values empirically attributed to anomalies conflict with the expectations these proposed applications suggest. In this paper, we take a closer look at the behavior of distribution densities through the lens of reparametrization and show that these quantities carry less meaningful information than previously thought, beyond estimation issues or the curse of dimensionality. We conclude that the use of these likelihoods for anomaly detection relies on strong and implicit hypotheses, and highlight the necessity of explicitly formulating these assumptions for reliable anomaly detection.
Accepted to the Special Issue "Probabilistic Methods for Deep Learning" of the Journal Entropy. 14 pages and 10 figures in main content, 4 pages of bibliography, and 2 pages in Appendix
References in corpus (24)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Array Programming with NumPy
- The NumPy array: a structure for efficient numerical computation
- NICE: Non-linear Independent Components Estimation
- A Unifying Review of Deep and Shallow Anomaly Detection
- Energy-based Out-of-distribution Detection
- Deep Anomaly Detection with Outlier Exposure
- Benchmarking Neural Network Robustness to Common Corruptions and Perturbations
- Towards a Critical Race Methodology in Algorithmic Fairness
- Flow++: Improving Flow-Based Generative Models with Variational Dequantization and Architecture Design
- Crypto-Nets: Neural Networks over Encrypted Data
- Does Object Recognition Work for Everyone?
- Can Autonomous Vehicles Identify, Recover From, and Adapt to Distribution Shifts?
- Contrastive Training for Improved Out-of-Distribution Detection
- Against Scale: Provocations and Resistances to Scale Thinking
- Why Normalizing Flows Fail to Detect Out-of-Distribution Data
- Hybrid Discriminative-Generative Training via Contrastive Learning
- Understanding Failures in Out-of-Distribution Detection with Deep Generative Models
- AI and the Everything in the Whole Wide World Benchmark
- Density of States Estimation for Out-of-Distribution Detection
- Further Analysis of Outlier Detection with Deep Generative Models
- Understanding the (un)interpretability of natural image distributions using generative models
- Deep Generative Models Strike Back! Improving Understanding and Evaluation in Light of Unmet Expectations for OoD Data
- Reparametrization Invariance in non-parametric Causal Discovery
Cited by in corpus (10)
- Anomaly Detection under Coordinate Transformations
- Deep machine learning for meteor monitoring: advances with transfer learning and gradient-weighted class activation mapping
- Graph Posterior Network: Bayesian Predictive Uncertainty for Node Classification
- Innovations Autoencoder and its Application in One-class Anomalous Sequence Detection
- Anomaly detection with flow-based fast calorimeter simulators
- Object classification on video data of meteors and meteor-like phenomena: algorithm and data
- Do We Really Need to Learn Representations from In-domain Data for Outlier Detection?
- New Methods and Datasets for Group Anomaly Detection From Fundamental Physics
- Autoencoding Under Normalization Constraints
- Entropic Issues in Likelihood-Based OOD Detection