Text Mining Utilities for Python 3
Project description
textmining3
Text Mining Utilities for Python 3
Free software: GNU General Public License v3
Documentation: https://textmining3.readthedocs.io.
Requires Python >= 3.6
Features
This package contains a variety of useful functions for text mining in Python 3.
It focuses on statistical text mining (i.e. the bag-of-words model) and makes it very easy to create a term-document matrix from a collection of documents. This matrix can then be read into a statistical package (R, MATLAB, etc.) for further analysis. The package also provides some useful utilities for finding collocations (i.e. significant two-word phrases), computing the edit distance between words, and chunking long documents up into smaller pieces.
The package has a large amount of curated data (stopwords, common names, an English dictionary with parts of speech and word frequencies) which allows the user to extract fairly sophisticated features from a document.
This package does NOT have any natural language processing capabilities such as part-of-speech tagging. Please see the Python NLTK for that sort of functionality (plus much, much more).
The original code and documentation is available in PyPI under the package name textmining. This package is a port to Python 3 and published in PyPI under the package name textmining3, and is based on the original.
Credits
The original textmining 1.0 package code was authored by Christian Peccei <cpeccei@hotmail.com>
This package was created with Cookiecutter and the audreyr/cookiecutter-pypackage project template.
History
1.1.0 (2018-13-19)
Add new feature to export DTM to pandas.DataFrame
1.0.2 (2018-12-19)
First port of textmining to Python 3
1.0.0 (2010-01-11)
Original release of textmining on PyPI (see https://pypi.org/project/textmining/1.0/)
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
File details
Details for the file textmining3-1.1.0.tar.gz
.
File metadata
- Download URL: textmining3-1.1.0.tar.gz
- Upload date:
- Size: 23.9 kB
- Tags: Source
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/1.12.1 pkginfo/1.4.2 requests/2.19.1 setuptools/40.4.3 requests-toolbelt/0.8.0 tqdm/4.26.0 CPython/3.7.0
File hashes
Algorithm | Hash digest | |
---|---|---|
SHA256 | 22cf971937a76f00722eadd0249b85bed9888cfaf57eaca238c8e55220c7bdb8 |
|
MD5 | 9a77cc5bc65751008772001e04c07517 |
|
BLAKE2b-256 | d37e78a5b991108302eb44b0a5347d274f9be6607fb41dd45c53e28244cce76f |
File details
Details for the file textmining3-1.1.0-py2.py3-none-any.whl
.
File metadata
- Download URL: textmining3-1.1.0-py2.py3-none-any.whl
- Upload date:
- Size: 1.9 MB
- Tags: Python 2, Python 3
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/1.12.1 pkginfo/1.4.2 requests/2.19.1 setuptools/40.4.3 requests-toolbelt/0.8.0 tqdm/4.26.0 CPython/3.7.0
File hashes
Algorithm | Hash digest | |
---|---|---|
SHA256 | e7a3c8ffc670caede8a6c1013c82082f3487e29fc3b652bdb9f56e1c66252f75 |
|
MD5 | 6eb6baf13e461a53b114ca27c7c34f27 |
|
BLAKE2b-256 | 14334d75039a7a9cd6bf07551cb2be035d43e6edacac5e2b5f3662d5f2343236 |