A Benchmark Study on Time Series Clustering
arXiv:2004.09546 · doi:10.1016/j.mlwa.2020.100001
Abstract
This paper presents the first time series clustering benchmark utilizing all time series datasets currently available in the University of California Riverside (UCR) archive -- the state of the art repository of time series data. Specifically, the benchmark examines eight popular clustering methods representing three categories of clustering algorithms (partitional, hierarchical and density-based) and three types of distance measures (Euclidean, dynamic time warping, and shape-based). We lay out six restrictions with special attention to making the benchmark as unbiased as possible. A phased evaluation approach was then designed for summarizing dataset-level assessment metrics and discussing the results. The benchmark study presented can be a useful reference for the research community on its own; and the dataset-level assessment metrics reported may be used for designing evaluation frameworks to answer different research questions.
Typos corrected, figures resolution changed
Cited by in corpus (7)
- A Review and Evaluation of Elastic Distance Functions for Time Series Clustering
- Unsupervised Deep Learning for IoT Time Series
- TCGAN: Convolutional Generative Adversarial Network for Time Series Classification and Clustering
- A Spatiotemporal Deep Neural Network for Fine-Grained Multi-Horizon Wind Prediction
- On the Use of Relative Validity Indices for Comparing Clustering Approaches
- Rock the KASBA: Blazingly Fast and Accurate Time Series Clustering
- Applications of Machine Learning in Pharmacogenomics: Clustering Plasma Concentration-Time Curves