Skip to main content

Text Mining Utilities for Python 3

Project description

textmining3

https://img.shields.io/pypi/v/textmining.svg https://img.shields.io/travis/djcomlab/textmining3.svg Documentation Status

Text Mining Utilities for Python 3

Features

This package contains a variety of useful functions for text mining in Python 3.

It focuses on statistical text mining (i.e. the bag-of-words model) and makes it very easy to create a term-document matrix from a collection of documents. This matrix can then be read into a statistical package (R, MATLAB, etc.) for further analysis. The package also provides some useful utilities for finding collocations (i.e. significant two-word phrases), computing the edit distance between words, and chunking long documents up into smaller pieces.

The package has a large amount of curated data (stopwords, common names, an English dictionary with parts of speech and word frequencies) which allows the user to extract fairly sophisticated features from a document.

This package does NOT have any natural language processing capabilities such as part-of-speech tagging. Please see the Python NLTK for that sort of functionality (plus much, much more).

The original code and documentation is available in PyPI under the package name textmining. This package is a port to Python 3 and published in PyPI under the package name textmining3, and is based on the original.

Credits

The original textmining 1.0 package code was authored by Christian Peccei <cpeccei@hotmail.com>

This package was created with Cookiecutter and the audreyr/cookiecutter-pypackage project template.

History

1.1.0 (2018-13-19)

  • Add new feature to export DTM to pandas.DataFrame

1.0.2 (2018-12-19)

  • First port of textmining to Python 3

1.0.0 (2010-01-11)

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

textmining3-1.1.0.tar.gz (23.9 kB view details)

Uploaded Source

Built Distribution

textmining3-1.1.0-py2.py3-none-any.whl (1.9 MB view details)

Uploaded Python 2 Python 3

File details

Details for the file textmining3-1.1.0.tar.gz.

File metadata

  • Download URL: textmining3-1.1.0.tar.gz
  • Upload date:
  • Size: 23.9 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/1.12.1 pkginfo/1.4.2 requests/2.19.1 setuptools/40.4.3 requests-toolbelt/0.8.0 tqdm/4.26.0 CPython/3.7.0

File hashes

Hashes for textmining3-1.1.0.tar.gz
Algorithm Hash digest
SHA256 22cf971937a76f00722eadd0249b85bed9888cfaf57eaca238c8e55220c7bdb8
MD5 9a77cc5bc65751008772001e04c07517
BLAKE2b-256 d37e78a5b991108302eb44b0a5347d274f9be6607fb41dd45c53e28244cce76f

See more details on using hashes here.

File details

Details for the file textmining3-1.1.0-py2.py3-none-any.whl.

File metadata

  • Download URL: textmining3-1.1.0-py2.py3-none-any.whl
  • Upload date:
  • Size: 1.9 MB
  • Tags: Python 2, Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/1.12.1 pkginfo/1.4.2 requests/2.19.1 setuptools/40.4.3 requests-toolbelt/0.8.0 tqdm/4.26.0 CPython/3.7.0

File hashes

Hashes for textmining3-1.1.0-py2.py3-none-any.whl
Algorithm Hash digest
SHA256 e7a3c8ffc670caede8a6c1013c82082f3487e29fc3b652bdb9f56e1c66252f75
MD5 6eb6baf13e461a53b114ca27c7c34f27
BLAKE2b-256 14334d75039a7a9cd6bf07551cb2be035d43e6edacac5e2b5f3662d5f2343236

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page