a library for doing approximate and phonetic matching of strings.
Project description
This is a simple fork of https://github.com/jamesturk/jellyfish package, no code change was done, only intention is to provide wheels for linux in pypi.
Jellyfish is a python library for doing approximate and phonetic matching of strings.
Written by James Turk <dev@jamesturk.net> and Michael Stephens.
See https://github.com/jamesturk/jellyfish/graphs/contributors for contributors.
See http://jellyfish.readthedocs.io for documentation.
Source is available at http://github.com/jamesturk/jellyfish.
Jellyfish >= 0.7 only supports Python 3, if you need Python 2 please use 0.6.x.
Included Algorithms
String comparison:
Levenshtein Distance
Damerau-Levenshtein Distance
Jaro Distance
Jaro-Winkler Distance
Match Rating Approach Comparison
Hamming Distance
Phonetic encoding:
American Soundex
Metaphone
NYSIIS (New York State Identification and Intelligence System)
Match Rating Codex
Example Usage
>>> import jellyfish >>> jellyfish.levenshtein_distance(u'jellyfish', u'smellyfish') 2 >>> jellyfish.jaro_distance(u'jellyfish', u'smellyfish') 0.89629629629629637 >>> jellyfish.damerau_levenshtein_distance(u'jellyfish', u'jellyfihs') 1
>>> jellyfish.metaphone(u'Jellyfish') 'JLFX' >>> jellyfish.soundex(u'Jellyfish') 'J412' >>> jellyfish.nysiis(u'Jellyfish') 'JALYF' >>> jellyfish.match_rating_codex(u'Jellyfish') 'JLLFSH'
Running Tests
If you are interested in contributing to Jellyfish, you may want to run tests locally. Jellyfish uses tox to run tests, which you can setup and run as follows:
pip install tox # cd jellyfish/ tox
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distributions
Hashes for jellyfish_wheel-0.7.0-cp37-cp37m-manylinux1_x86_64.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | d7534548bddeef42eebe3ec38c9fdbf1abb45b0eb3fa97e3f8c294d6cb96b21b |
|
MD5 | a95dcb1f52ffa6c427c5f8d2aa94deed |
|
BLAKE2b-256 | ee53900a04c74f565de89a8b9dff6b19e20de809e9a169fd89d7f09731c521e2 |
Hashes for jellyfish_wheel-0.7.0-cp36-cp36m-manylinux1_x86_64.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | 9a090d2e2d48ad4c480d9b7a27ea8d8c27ace08e71029db70479f3236fa1e357 |
|
MD5 | 7882b29a7a07080dabbf8b0cb7cfd2a5 |
|
BLAKE2b-256 | 5b600196173a0ddcf8f83ef19dd556385bb9153090c735cbc7ba8bb0ef16d4ee |
Hashes for jellyfish_wheel-0.7.0-cp35-cp35m-manylinux1_x86_64.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | 08c0cceaf4c032cf4b8dba2692319a2984f88201674d70f42862a2c985241022 |
|
MD5 | 34f4bb0aa75052535531b13eaa7ac28d |
|
BLAKE2b-256 | 5103328b3849fd2074aa9d45a9508b66f6c89c5f56b4a84ee8bb5b89eb728c83 |
Hashes for jellyfish_wheel-0.7.0-cp34-cp34m-manylinux1_x86_64.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | 4e88592dda21e017a2ee43a9e284f2447a8d75b1c3fe9d74731c351348432646 |
|
MD5 | 13b5ffddaceb889c5a57fe760989be23 |
|
BLAKE2b-256 | e13bb0dfad375896f378855ef1bce8b790c525b3c403fdac9265cefc844a1a13 |