Skip to main content

No project description provided

Project description

NLP Primitives

CircleCI

nlp_primitives is a Python library with Natural Language Processing Primitives, intended for use with Featuretools.

nlp_primitives allows you to make use of text data in your machine learning pipeline in the same pipeline as the rest of your data.

Install

pip install 'featuretools[nlp_primitives]'

Demos

Calculating Features

With nlp_primitives primtives in featuretools, this is how to calculate the same feature.

from featuretools.nlp_primitives import PolarityScore

data = ["hello, this is a new featuretools library",
        "this will add new natural language primitives",
        "we hope you like it!"]

pol = PolarityScore()
pol(data)
0    0.365
1    0.385
2    1.000
dtype: float64

Combining Primitives

In featuretools, this is how to combine nlp_primitives primitives with built-in or other installed primitives.

import featuretools as ft
from featuretools.nlp_primitives import TitleWordCount
from featuretools.primitives import Mean

entityset = ft.demo.load_retail()
feature_matrix, features = ft.dfs(entityset=entityset, target_entity='products', agg_primitives=[Mean], trans_primitives=[TitleWordCount])

feature_matrix.head(5)
           MEAN(order_products.quantity)  MEAN(order_products.unit_price)  MEAN(order_products.total)  TITLE_WORD_COUNT(description)
product_id
10002                         16.795918                          1.402500                   23.556276                           3.0
10080                         13.857143                          0.679643                    8.989357                           3.0
10120                          6.620690                          0.346500                    2.294069                           2.0
10123C                         1.666667                          1.072500                    1.787500                           3.0
10124A                           3.2000                            0.6930                      2.2176                           5.0

Feature Labs

Featuretools

NLP Primitives is an open source project created by Feature Labs. To see the other open source projects we're working on visit Feature Labs Open Source. If building impactful data science pipelines is important to you or your business, please get in touch.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

nlp_primitives-0.3.0.tar.gz (12.9 kB view details)

Uploaded Source

Built Distribution

nlp_primitives-0.3.0-py3-none-any.whl (22.6 kB view details)

Uploaded Python 3

File details

Details for the file nlp_primitives-0.3.0.tar.gz.

File metadata

  • Download URL: nlp_primitives-0.3.0.tar.gz
  • Upload date:
  • Size: 12.9 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.1.1 pkginfo/1.5.0.1 requests/2.23.0 setuptools/41.2.0 requests-toolbelt/0.9.1 tqdm/4.46.0 CPython/3.7.7

File hashes

Hashes for nlp_primitives-0.3.0.tar.gz
Algorithm Hash digest
SHA256 6a46fd236534bd545e6407b88b3e9cb2542730f9557f7d5bd25a58e0b76e5320
MD5 049816d4aa6a9a0b5b827a75d4585a8d
BLAKE2b-256 055d7d9eb9e61faf7d928b811b428f94e899b36dcebcc7c416d2cd6c8a81d082

See more details on using hashes here.

File details

Details for the file nlp_primitives-0.3.0-py3-none-any.whl.

File metadata

  • Download URL: nlp_primitives-0.3.0-py3-none-any.whl
  • Upload date:
  • Size: 22.6 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.1.1 pkginfo/1.5.0.1 requests/2.23.0 setuptools/41.2.0 requests-toolbelt/0.9.1 tqdm/4.46.0 CPython/3.7.7

File hashes

Hashes for nlp_primitives-0.3.0-py3-none-any.whl
Algorithm Hash digest
SHA256 c2eda904b165a154648c062f269dbaf244f640b25490469a8f3185556a959b27
MD5 7d4254c7c8a7186d8e9e7b8b4a578623
BLAKE2b-256 f7a01f782243bac125e366233376cd0e6c18b8f33a448254bfdc5b11dd7c7bf7

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page