Skip to main content

A tool for building feature stores - Transform your raw data into beautiful features.

Project description

Butterfree

A tool for building feature stores. Transform your raw data into beautiful features.

Release Python Version License Code style: black

Source Downloads Page Installation Command
PyPi PyPi Downloads Link pip install butterfree

Build status

Develop Stable Documentation Sonar
Test Publish Documentation Status Quality Gate Status

Made with :heart: by the MLOps team from QuintoAndar

This library supports Python version 3.7+ and meant to provide tools for building ETL pipelines for Feature Stores using Apache Spark.

The library is centered on the following concetps:

  • ETL: central framework to create data pipelines. Spark-based Extract, Transform and Load modules ready to use.
  • Declarative Feature Engineering: care about what you want to compute and not how to code it.
  • Feature Store Modeling: the library easily provides everything you need to process and load data to your Feature Store.

To understand the main concepts of Feature Store modeling and library main features you can check Butterfree's Documentation, which is hosted by Read the Docs.

To learn how to use Butterfree in practice, see Butterfree's notebook examples

Requirements and Installation

Butterfree depends on Python 3.7+ and it is Spark 3.0 ready :heavy_check_mark:

Python Package Index hosts reference to a pip-installable module of this library, using it is as straightforward as including it on your project's requirements.

pip install butterfree

Or after listing butterfree in your requirements.txt file:

pip install -r requirements.txt

Dev Package are available for testing using the .devN versions of the Butterfree on PyPi.

License

Apache License 2.0

Contributing

All contributions are welcome! Feel free to open Pull Requests. Check the development and contributing guidelines described here.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

butterfree-1.2.0.dev7.tar.gz (61.1 kB view details)

Uploaded Source

Built Distribution

butterfree-1.2.0.dev7-py3-none-any.whl (104.1 kB view details)

Uploaded Python 3

File details

Details for the file butterfree-1.2.0.dev7.tar.gz.

File metadata

  • Download URL: butterfree-1.2.0.dev7.tar.gz
  • Upload date:
  • Size: 61.1 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.1.1 pkginfo/1.7.0 requests/2.25.1 setuptools/41.6.0 requests-toolbelt/0.9.1 tqdm/4.60.0 CPython/3.7.9

File hashes

Hashes for butterfree-1.2.0.dev7.tar.gz
Algorithm Hash digest
SHA256 c7e5a10906da4962b0949a2623bcf9a54943b7bfc2dd1d1356d4c26510a00b53
MD5 f419bd52421655f94c8f0a30f382a094
BLAKE2b-256 be1b8d3ea77517dbe2f2cef51436173cab3d70cae41973028df2bbf824eca545

See more details on using hashes here.

File details

Details for the file butterfree-1.2.0.dev7-py3-none-any.whl.

File metadata

  • Download URL: butterfree-1.2.0.dev7-py3-none-any.whl
  • Upload date:
  • Size: 104.1 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/3.1.1 pkginfo/1.7.0 requests/2.25.1 setuptools/41.6.0 requests-toolbelt/0.9.1 tqdm/4.60.0 CPython/3.7.9

File hashes

Hashes for butterfree-1.2.0.dev7-py3-none-any.whl
Algorithm Hash digest
SHA256 1d95d0ad75c2cc44d304f5958f67108204ec7d216aac35f0565e2dac5e340fff
MD5 d2c866d1102decfedb8c0d1b245b0862
BLAKE2b-256 61ed51f94b04673280f53275bd8cb065d0744c171e8c458d0718256130f6ac12

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page