Skip to main content

Python library for end to end Entity Matching.

Project description


Entity matching (EM) is a critical problem and will become even more critical. Lot of EM research, but has focused mostly on matching algorithms. Focus mostly on maximizing some performance factors such as accuracy, time, money, etc. Currently there is very little help for users. For many of these steps, there is little or no work, so no solution. Even when there are solutions, there may be no tools. Even when there are tools, user is faced with a wide variety of tools, one for each step, so user is left trying to move among the tools, stitching them together, deciphering their different data formats, import/export commands. too cumbersome and difficult. So often users just give up and roll their own ad-hoc solutions instead. This is obviously a major bottleneck that prevents the deployment of matching in practice. To solve the above problems, py_entitymatching aims to provide a set of commands that help the user to

  • Come up with a EM workflow
  • Iterate and debugmatcher the workflow
  • Deploy it in production

The package is free, open-source, and BSD-licensed.


py_entitymatching has been tested on Linux, OS X and Windows.

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Files for py-entitymatching, version 0.0.0
Filename, size File type Python version Upload date Hashes
Filename, size py_entitymatching-0.0.0.tar.gz (463.0 kB) File type Source Python version None Upload date Hashes View

Supported by

Pingdom Pingdom Monitoring Google Google Object Storage and Download Analytics Sentry Sentry Error logging AWS AWS Cloud computing DataDog DataDog Monitoring Fastly Fastly CDN DigiCert DigiCert EV certificate StatusPage StatusPage Status page