A scikit-based Python environment for performing multi-label classification
arXiv:1702.01460
Abstract
scikit-multilearn is a Python library for performing multi-label classification. The library is compatible with the scikit/scipy ecosystem and uses sparse matrices for all internal operations. It provides native Python implementations of popular multi-label classification methods alongside a novel framework for label space partitioning and division. It includes modern algorithm adaptation methods, network-based label space division approaches, which extracts label dependency information and multi-label embedding classifiers. It provides python wrapped access to the extensive multi-label method stack from Java libraries and makes it possible to extend deep learning single-label methods for multi-label tasks. The library allows multi-label stratification and data set management. The implementation is more efficient in problem transformation than other established libraries, has good test coverage and follows PEP8. Source code and documentation can be downloaded from http://scikit.ml and also via pip. The library follows BSD licensing scheme.
References in corpus (2)
Cited by in corpus (14)
- Classifier Chains: A Review and Perspectives
- SGM: Sequence Generation Model for Multi-label Classification
- Multi-Label Classification Neural Networks with Hard Logical Constraints
- Predicting Different Types of Subtle Toxicity in Unhealthy Online Conversations
- Classifications of Skull Fractures using CT Scan Images via CNN with Lazy Learning Approach
- NCoRE: Neural Counterfactual Representation Learning for Combinations of Treatments
- Prototypical Networks for Multi-Label Learning
- Hierarchical Text Classification of Urdu News using Deep Neural Network
- LNEMLC: Label Network Embeddings for Multi-Label Classification
- CliniQG4QA: Generating Diverse Questions for Domain Adaptation of Clinical Question Answering
- A Unified Framework for Multiclass and Multilabel Support Vector Machines
- Novel split quality measures for stratified multilabel Cross Validation with application to large and sparse gene ontology datasets
- A PAC-Bayesian Perspective on Structured Prediction with Implicit Loss Embeddings
- MLPSVM:A new parallel support vector machine to multi-label learning