mldr.resampling: Efficient Reference Implementations of Multilabel Resampling Algorithms
arXiv:2305.17152 · doi:10.1016/j.neucom.2023.126806
Abstract
Resampling algorithms are a useful approach to deal with imbalanced learning in multilabel scenarios. These methods have to deal with singularities in the multilabel data, such as the occurrence of frequent and infrequent labels in the same instance. Implementations of these methods are sometimes limited to the pseudocode provided by their authors in a paper. This Original Software Publication presents mldr.resampling, a software package that provides reference implementations for eleven multilabel resampling methods, with an emphasis on efficiency since these algorithms are usually time-consuming.
References in corpus (5)
- SMOTE: Synthetic Minority Over-sampling Technique
- Scalable Multi-Output Label Prediction: From Classifier Chains to Classifier Trellises
- Dealing with Difficult Minority Labels in Imbalanced Mutilabel Data Sets
- Tips, guidelines and tools for managing multi-label datasets: the mldr.datasets R package and the Cometa data repository
- A snapshot on nonstandard supervised learning problems: taxonomy, relationships and methods