scikit-multilearn

Scikit-multilearn is a BSD-licensed library for multi-label classification that is built on top of the well-known scikit-learn ecosystem.

These details have not been verified by PyPI

Project links

Homepage

Project description

# scikit-multilearn

[![PyPI version](https://badge.fury.io/py/scikit-multilearn.svg)](https://badge.fury.io/py/scikit-multilearn)
[![License](https://img.shields.io/badge/License-BSD%202--Clause-orange.svg)](https://opensource.org/licenses/BSD-2-Clause)
[![Build Status Linux and OSX](https://travis-ci.org/scikit-multilearn/scikit-multilearn.svg?branch=master)](https://travis-ci.org/scikit-multilearn/scikit-multilearn)
[![Build Status Windows](https://ci.appveyor.com/api/projects/status/vd4k18u1lp5btaql/branch/master?svg=true)](https://ci.appveyor.com/project/niedakh/scikit-multilearn/branch/master)

__scikit-multilearn__ is a Python module capable of performing multi-label
learning tasks. It is built on-top of various scientific Python packages
([numpy](http://www.numpy.org/), [scipy](https://www.scipy.org/)) and
follows a similar API to that of [scikit-learn](http://scikit-learn.org/).

- __Website:__ [scikit.ml](http://scikit.ml)
- __Documentation:__ [scikit-multilearn Documentation](http://scikit.ml/api/skmultilearn.html)

## Features

- __Native Python implementation.__ A native Python implementation for a variety of multi-label classification algorithms. To see the list of all supported classifiers, check this [link](http://scikit.ml/#classifiers).

- __Interface to Meka.__ A Meka wrapper class is implemented for reference purposes and integration. This provides access to all methods available in MEKA, MULAN, and WEKA — the reference standard in the field.

- __Builds upon giants!__ Team-up with the power of numpy and scikit. You can use scikit-learn's base classifiers as scikit-multilearn's classifiers. In addition, the two packages follow a similar API.

## Dependencies

In most cases you will want to follow the requirements defined in the requirements/*.txt files in the package.

### Base dependencies
```
scipy
numpy
future
scikit-learn
liac-arff # for loading ARFF files
requests # for dataset module
networkx # for networkX base community detection clusterers
python-louvain # for networkX base community detection clusterers
keras
```

### GPL-incurring dependencies for two clusterers
```
python-igraph # for igraph library based clusterers
python-graphtool # for graphtool base clusterers
```

Note: Installing graphtool is complicated, please see: [graphtool install instructions](https://git.skewed.de/count0/graph-tool/wikis/installation-instructions)

## Installation

To install scikit-multilearn, simply type the following command:

```bash
$ pip install scikit-multilearn
```

This will install the latest release from the Python package index. If you
wish to install the bleeding-edge version, then clone this repository and
run `setup.py`:

```bash
$ git clone https://github.com/scikit-multilearn/scikit-multilearn.git
$ cd scikit-multilearn
$ python setup.py
```

## Basic Usage

Before proceeding to classification, this library assumes that you have
a dataset with the following matrices:

- `x_train`, `x_test`: training and test feature matrices of size `(n_samples, n_features)`
- `y_train`, `y_test`: training and test label matrices of size `(n_samples, n_labels)`

Suppose we wanted to use a problem-transformation method called Binary
Relevance, which treats each label as a separate single-label classification
problem, to a Support-vector machine (SVM) classifier, we simply perform
the following tasks:

```python
# Import BinaryRelevance from skmultilearn
from skmultilearn.problem_transform import BinaryRelevance

# Import SVC classifier from sklearn
from sklearn.svm import SVC

# Setup the classifier
classifier = BinaryRelevance(classifier=SVC(), require_dense=[False,True])

# Train
classifier.fit(X_train, y_train)

# Predict
y_pred = classifier.predict(X_test)
```

More examples and use-cases can be seen in the
[documentation](http://scikit.ml/api/classify.html). For using the MEKA
wrapper, check this [link](http://scikit.ml/api/meka.html#mekawrapper).

## Contributing

This project is open for contributions. Here are some of the ways for
you to contribute:

- Bug reports/fix
- Features requests
- Use-case demonstrations
- Documentation updates

In case you want to implement your own multi-label classifier, please
read our [Developer's Guide](http://scikit.ml/api/base.html) to help
you integrate your implementation in our API.

To make a contribution, just fork this repository, push the changes
in your fork, open up an issue, and make a Pull Request!

We're also available in Slack! Just go to our [slack group](https://scikit-ml.slack.com/).

## Cite

If you used scikit-multilearn in your research or project, please
cite [our work](https://arxiv.org/abs/1702.01460):

```bibtex
@ARTICLE{2017arXiv170201460S,
author = {{Szyma{\'n}ski}, P. and {Kajdanowicz}, T.},
title = "{A scikit-based Python environment for performing multi-label classification}",
journal = {ArXiv e-prints},
archivePrefix = "arXiv",
eprint = {1702.01460},
year = 2017,
month = feb
}
```

Project details

These details have not been verified by PyPI

Project links

Homepage

Release history Release notifications | RSS feed

This version

0.2.0

Dec 10, 2018

0.1.0

Sep 3, 2018

0.0.5

Feb 25, 2017

0.0.4

Feb 10, 2017

0.0.3

Jun 3, 2016

0.0.2

Jun 3, 2016

0.0.1

Nov 27, 2014

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

scikit-multilearn-0.2.0.linux-x86_64.tar.gz (113.7 kB view details)

Uploaded Dec 10, 2018 Source

Built Distributions

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

scikit_multilearn-0.2.0-py3-none-any.whl (89.4 kB view details)

Uploaded Dec 10, 2018 Python 3

scikit_multilearn-0.2.0-py2-none-any.whl (89.4 kB view details)

Uploaded Dec 10, 2018 Python 2

File details

Details for the file scikit-multilearn-0.2.0.linux-x86_64.tar.gz.

File metadata

Download URL: scikit-multilearn-0.2.0.linux-x86_64.tar.gz
Upload date: Dec 10, 2018
Size: 113.7 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/1.8.1 pkginfo/1.2.1 requests/2.18.1 setuptools/36.2.7 requests-toolbelt/0.8.0 clint/0.5.1 CPython/3.6.3 Linux/4.13.0-46-generic

File hashes

Hashes for scikit-multilearn-0.2.0.linux-x86_64.tar.gz
Algorithm	Hash digest
SHA256	`3179fed29b1492f6a69600696c23045b9f494d2b89d1796a8bdc43ccbb33712b`
MD5	`c23e41337785bfb7f63c637e3d5c6cb9`
BLAKE2b-256	`fc574c8951d3613c1cd569910bc3ddd5b3a755ad383297a12004eb2d61eefc06`

See more details on using hashes here.

File details

Details for the file scikit_multilearn-0.2.0-py3-none-any.whl.

File metadata

Download URL: scikit_multilearn-0.2.0-py3-none-any.whl
Upload date: Dec 10, 2018
Size: 89.4 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/1.8.1 pkginfo/1.2.1 requests/2.18.1 setuptools/36.2.7 requests-toolbelt/0.8.0 clint/0.5.1 CPython/3.6.3 Linux/4.13.0-46-generic

File hashes

Hashes for scikit_multilearn-0.2.0-py3-none-any.whl
Algorithm	Hash digest
SHA256	`068c652f22704a084ca252d05d21a655e7c9b248d0a4543847b74de5fca2b3f0`
MD5	`d178b98b1a0320135d6b80f14058606e`
BLAKE2b-256	`bb1fe6ff649c72a1cdf2c7a1d31eb21705110ce1c5d3e7e26b2cc300e1637272`

See more details on using hashes here.

File details

Details for the file scikit_multilearn-0.2.0-py2-none-any.whl.

File metadata

Download URL: scikit_multilearn-0.2.0-py2-none-any.whl
Upload date: Dec 10, 2018
Size: 89.4 kB
Tags: Python 2
Uploaded using Trusted Publishing? No
Uploaded via: twine/1.8.1 pkginfo/1.2.1 requests/2.18.1 setuptools/36.2.7 requests-toolbelt/0.8.0 clint/0.5.1 CPython/3.6.3 Linux/4.13.0-46-generic

File hashes

Hashes for scikit_multilearn-0.2.0-py2-none-any.whl
Algorithm	Hash digest
SHA256	`0a389600a6797db6567f2f6ca1d0dca30bebfaaa73f75de62d7ae40f8f03d4fb`
MD5	`fc1a5adba295e7e7c0642b2a88d76780`
BLAKE2b-256	`21829a5a40ac8bcf4be9662728df35d7e0a1f47454416e4d4480b8d0a1dc1a7b`

See more details on using hashes here.

scikit-multilearn 0.2.0

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distributions

File details

File metadata

File hashes

File details

File metadata

File hashes

File details

File metadata

File hashes