Directional Statistics in Machine Learning: a Brief Review
arXiv:1605.00316
Abstract
The modern data analyst must cope with data encoded in various forms, vectors, matrices, strings, graphs, or more. Consequently, statistical and machine learning models tailored to different data encodings are important. We focus on data encoded as normalized vectors, so that their "direction" is more important than their magnitude. Specifically, we consider high-dimensional vectors that lie either on the surface of the unit hypersphere or on the real projective plane. For such data, we briefly review common mathematical models prevalent in machine learning, while also outlining some technical aspects, software, applications, and open mathematical challenges.
12 pages, slightly modified version of submitted book chapter
Cited by in corpus (10)
- Weakly-Supervised Neural Text Classification
- Statistical and Topological Properties of Sliced Probability Divergences
- The Power Spherical distribution
- Remarks on Optimal Scores for Speaker Recognition
- On the Spherical Dirichlet Distribution: Corrections and Results
- A Consistently Oriented Basis for Eigenanalysis
- Spherical Sliced-Wasserstein
- High-Dimensional Bayesian Optimization via Nested Riemannian Manifolds
- 6D Camera Relocalization in Ambiguous Scenes via Continuous Multimodal Inference
- Improving Relational Regularized Autoencoders with Spherical Sliced Fused Gromov Wasserstein