Skip to main content

Playbooks for data. Open, process and save table based data.

Project description

Data Playbook

:book: Playbooks for data. Open, process and save table based data.

Automate repetitive tasks on table based data. Include various input and output tasks. Can be extended with custom modules.

Install: pip install dataplaybook

Use: dataplaybook playbook.yaml

Playbook structure

The playbook.yaml file allows you to load additional modules (containing tasks) and specify the tasks to execute in sequence, with all their parameters.

The tasks to perform typically follow the the structure of read, process, write.

Example yaml: (please note yaml is case sensitive)

modules: [list, of, modules]

tasks:
  - task: *name
    tables: # List of tables used by this task
    target: # Target table name of this function
    debug*: True/False # default: False
    # task specific properties, refer to each task

Tasks

Tasks are implemented as simple Python functions and the modules can be found in the dataplaybook/tasks folder.

Default tasks

  • drop
  • extend
  • filter
  • fuzzy_match (pip install fuzzywuzzy)
  • print
  • replace
  • unique
  • vlookup

Module io_xlsx (loaded by default)

  • read_excel
  • write_excel

Module io_misc (loaded by default)

  • read_tab_delim
  • read_text_regex
  • wget
  • write_csv

Module io_mongo (uses pymongo)

  • read_mongo
  • write_mongo
  • columns_to_list
  • list_to_columns

Module io_pdf (requires pdftotext)

  • read_pdf_pages
  • read_pdf_files

Module io_xml

Module ietf

Module gis

Module fnb

Install development version

  1. Clone the repo
  2. pip install <path> -e

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

dataplaybook-0.2.2.tar.gz (21.6 kB view details)

Uploaded Source

File details

Details for the file dataplaybook-0.2.2.tar.gz.

File metadata

File hashes

Hashes for dataplaybook-0.2.2.tar.gz
Algorithm Hash digest
SHA256 82d50cfa32d13ee53916fe4fa1541b7d0f2df4ca7b4d274695c2b322c8cd107f
MD5 08c80847e0db615be20e274497558dc0
BLAKE2b-256 a33b4665622487665a02bd1e285a28ef0d727e2afc90cdebd43aa0c4fe18ed17

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page