Skip to main content

A tool for bulk river creation

Project description

Azure Blob Storage batch rivers creation CLI Tool

A command-line interface (CLI) tool to interact with Azure Blob Storage and manage rivers in Rivery. This tool allows you to configure Azure and Rivery credentials, generate river templates based on blob names, populate rivers, delete rivers, and list source configurations.

Table of Contents

Prerequisites

  • Python 3.6 or higher
  • Azure Blob Storage account
  • Rivery account and API token
  • Virtual Environment: It is recommended to use a virtual environment to manage dependencies.

Installation

  1. Create a Virtual Environment

    It is recommended to create a virtual environment to isolate the dependencies.

    python3 -m venv venv
    source venv/bin/activate
    
  2. Install the CLI Tool

    Install the multiriver package using pip:

    pip install multiriver
    

    (Ensure that the package multiriver is available in the Python Package Index or replace this command with the appropriate installation method if installing from a local source.)

Configuration

Before using the CLI tool, you need to configure your Azure and Rivery credentials, as well as source settings.

Configure Credentials

Use the configure_creds command to set up your Azure Storage and Rivery API credentials.

multiriver configure-creds \
    --account-name YOUR_AZURE_ACCOUNT_NAME \
    --account-key YOUR_AZURE_ACCOUNT_KEY \
    --rivery-api-token YOUR_RIVERY_API_TOKEN

Options:

  • --account-name (optional, but mandatory for the connection): Azure Storage account name.
  • --account-key (optional, but mandatory for the connection): Azure Storage account key.
  • --connection-string (optional): Azure Storage connection string.
  • --sas-token (optional): Shared Access Signature (SAS) token.
  • --rivery-api-token (optional, but mandatory for the connection): Rivery API token.
  • --rivery-host (optional): Rivery host URL (default: https://console.rivery.io).

Note: You must provide at least one Azure credential and the Rivery API token.

Configure Source

Use the configure-source command to set up source settings for Azure Blob Storage. This command will generate river templates based on the blobs in the specified container and store them together with the config under ~/.rivery/source_config.

multiriver configure-source \
    --container-name YOUR_CONTAINER_NAME \
    --template-river-id TEMPLATE_RIVER_ID \
    --filename-template FILENAME_TEMPLATE \
    --group-id GROUP_ID \
    --cron-schedule CRON_SCHEDULE

Options:

  • --container-name (optional): Name of the Azure Blob Storage container.
  • it�� (optional): ID of the base template river in Rivery.
  • --filename-template (optional): Template used for generating river names.
  • --group-id (required): ID of the group to attach the rivers to in Rivery.
  • --cron-schedule (optional): Cron expression to schedule new rivers.

Note: You must provide the --group-id option. If --filename-template is provided, river templates will be generated based on the blob names.

Commands

populate-rivers

Populate rivers in Rivery based on the generated templates.

multiriver populate-rivers --group-id GROUP_ID

Options:

  • --group-id (required): ID of the group to attach the rivers to in Rivery.

delete-rivers

Delete all rivers associated with a specific group in Rivery.

multiriver delete-rivers --group-id GROUP_ID

Options:

  • --group-id (required): ID of the group whose rivers will be deleted.

list-source-configs

List all source configurations that have been set up.

multiriver list-source-configs

Usage Examples

Example Workflow

  1. Create a Virtual Environment

    python3 -m venv venv
    source venv/bin/activate
    
  2. Install the CLI Tool

    pip install multiriver
    
  3. Configure Credentials

    multiriver configure_creds \
        --account-name myazureaccount \
        --account-key myazurekey \
        --rivery-api-token myriverytoken
    
  4. Configure Source

    multiriver configure-source \
        --container-name mycontainer \
        --template-river-id 62b075d34c86b10010ddf473 \
        --filename-template "{entity_name}_data.csv" \
        --group-id 6720e1592f775cb9fcdbf026 \
        --cron-schedule "0 0 * * *"
    
  5. Populate Rivers

    multiriver populate-rivers --group-id 6720e1592f775cb9fcdbf026
    

    This command will:

    • Create new rivers in Rivery using the templates.
    • Generate mapping for every river.
    • Schedule the rivers if a cron schedule was provided.
  6. Delete Rivers

    If you need to delete all rivers associated with the group:

    multiriver delete-rivers --group-id 6720e1592f775cb9fcdbf026
    
  7. List Source Configurations

    To view all source configurations:

    multiriver list-source-configs
    

License

This project is licensed under the MIT License - see the LICENSE file for details.

Contact

For any questions or issues, please contact Rivery support.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

multiriver-1.1.8.tar.gz (16.2 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

multiriver-1.1.8-py3-none-any.whl (16.0 kB view details)

Uploaded Python 3

File details

Details for the file multiriver-1.1.8.tar.gz.

File metadata

  • Download URL: multiriver-1.1.8.tar.gz
  • Upload date:
  • Size: 16.2 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/5.1.1 CPython/3.8.8

File hashes

Hashes for multiriver-1.1.8.tar.gz
Algorithm Hash digest
SHA256 bae7eadd08eee0d219926dc69c2c8dd1ba30804337537e7307317aaef805baa2
MD5 791932a3f847e6ff767a60e6958de7c7
BLAKE2b-256 f61d3763436f73c98898f03b935a2abbcfab57d9a6760cc74ad099f55bcdf984

See more details on using hashes here.

File details

Details for the file multiriver-1.1.8-py3-none-any.whl.

File metadata

  • Download URL: multiriver-1.1.8-py3-none-any.whl
  • Upload date:
  • Size: 16.0 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/5.1.1 CPython/3.8.8

File hashes

Hashes for multiriver-1.1.8-py3-none-any.whl
Algorithm Hash digest
SHA256 290537c92865bf4220ae254579655d2fd21281c61974b6d0c8448b552174b319
MD5 320241276230afda9e3c0156db5fbdf6
BLAKE2b-256 08f1797f7d08973be8f4d702d20fa744fcda831964721586846c61f70e7cc04d

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page