Skip to main content

Tools to help manage external data for Harness CCM

Project description

harness-ccm-external-data

PyPI - Version Docker Image Version

tools to help manage ingesting external data in harness ccm

the project is split into several parts:

  • creating raw data
  • converting external data formats to a compatible format for harness (focus)
  • uploading the converted data to the harness platform

loading data

when loading in a billing export we can apply a few modifications of the data to prepare it for ingestion into harness.

first, we convert any non-focus fields to their focus equivalent. this is done by providing a map of focus fields to their corresponding non-focus fields.

mapping = {
    "BillingAccountId": "Organization ID",
    "BillingAccountName": "Organization Name",
    ...
}

if only a subset of fields need remapping you can specify only those which need changed.

next we create a Focus object, specifying the platform, local billing export file, field mappings (if needed) and any data modifications needed:

from harness_ccm_external_data import Focus

my_data = Focus(
    # name of the provider where the billing data came from
    provider="CloudABC",
    # name of this particular data source from the above provider
    data_source="ABC Payer Account 1",
    # csv with focus data
    filename="abc_billing_export.csv",
    # focus to non focus mappings (if they exist)
    mapping={
        "BillingAccountId": "Organization ID",
        "BillingAccountName": "Organization Name",
        ...
    },
    # skip the first n rows of the billing data
    skip_rows=100,
    # you can also specify specific rows to skip
    # skip_rows=[0, 2, 4, 6, 8],
    # apply a multiplier to the cost (account for discounts not shown in the export?)
    cost_multiplier=0.95
    # if the csv is in a non-standard format
    separator=";"
    # apply a function to any column value
    converters={
        "ChargeCategory": lambda x: lower(x)
    },
    # fill in missing required fields with static data if needed
    additional_columns={
        "ConsumedQuantity": 1,
    },
    # for data upload to harness
    harness_account_id=getenv("HARNESS_ACCOUNT_ID"),
    harness_platform_api_key=getenv("HARNESS_PLATFORM_API_KEY"),
)

now we can render the data to the harness platform format to be uploaded by hand in the UI:

my_data.render_file("harness_focus_my_billing_export.csv")

building data

you can also build cost data on the fly by building a dataframe and passing that to the constructor:

from harness_ccm_external_data import Focus

data = [
    [
        1234567890123,
        "SunBird",
        "2025-6-01 00:00:00",
        "2025-5-01 00:00:00",
        "Usage",
        "2024-09-18 22:00:00",
        "2024-09-18 23:00:00",
        2.0,
        0.0,
        "AWS",
        "arn:ats:sqs:us-test-2:347410479675:mibelllmel-i-032l64f2065481b12",
        "US West (Oregon)",
        "Amazon Simple Queue Service",
        51738928782,
        "G95FST5FTYV3JSRX",
        "Atlas Nimbus",
    ],
    [
        1234567890124,
        "SunBird2",
        "2025-6-01 00:00:00",
        "2025-5-01 00:00:00",
        "Usage",
        "2024-09-30 22:00:00",
        "2024-09-30 23:00:00",
        0.00200749,
        0.0,
        "AWS",
        "arn:ats:emastilmoalfamanling:us-test-2:586597448978:moalfamanler/app/tungsten-lonbmuenle-amf/l365455f461l4e4a",
        "US West (Oregon)",
        "Elastic Load Balancing",
        43883916739,
        "2ETY8Y426S4237JU",
        "Zenith Eclipse",
        '{"application": "BrightLensMatrix", "environment": "dev", "business_unit": "ViennaAI"}',
    ],
]

my_data = Focus(
    provider="CloudABC",
    data_source="ABC Payer Account 1",
    source=Focus.create_dataset(data),
)

the data defined should have the harness focus feilds as defined in HARNESS_FIELDS

uploading data

uploading the data to harness is as simple as executing the upload function, there is no need to render the data to a file before doing so:

my_data.upload()

this will auto-detect the invoice period from the data, upload it, and trigger ingestion

docker

there is a docker image available to enable running the automation via docker or a plugin in a harness pipeline:

docker run --rm -it \
  -v ${PWD}/focus_sample.csv:/focus_sample.csv \
  -v ${PWD}:/output \
  -e CSV_FILE=/focus_sample.csv \
  -e PROVIDER=CloudABC \
  -e DATA_SOURCE="ABC Payer Account 1" \
  -e RENDER_FILE=/output/docker_focus.csv # optional \
  -e UPLOAD=true # optional \
  harnesscommunity/harness-ccm-external-data

drone plugin

the container can also be used as a drone/harness plugin:

- step:
    type: Plugin
    name: upload
    identifier: upload
    spec:
        connectorRef: account.buildfarm_container_registry_cloud
        image: harnesscommunity/harness-ccm-external-data
        settings:
            PROVIDER: CloudABC
            DATA_SOURCE: ABC Payer Account 1
            CSV_FILE: /harness/focus_sample.csv
            HARNESS_ACCOUNT_ID: <+account.identifier>
            HARNESS_PLATFORM_API_KEY: <+secrets.getValue("account.account_admin")>
            UPLOAD: "true" # optional
            RENDER_FILE: /harness/harness_focus_sample.csv # optional

modules

there are patterns provided for extracting, transforming, and loading external data into harness under the modules folder:

  • aws: s3+lambda function

data loading settings

  • RENDER_FILE: file path to render harness-focus data to
  • PROVIDER:
  • CSV_FILE:
  • MAPPING:
  • SKIP_ROWS:
  • COST_MULTIPLIER:
  • VALIDATE:

development

pull the example focus csv: curl -LO https://raw.githubusercontent.com/FinOps-Open-Cost-and-Usage-Spec/FOCUS-Sample-Data/refs/heads/main/FOCUS-1.0/focus_sample.csv

install poetry

testing: make test

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

harness_ccm_external_data-0.1.5.tar.gz (7.7 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

harness_ccm_external_data-0.1.5-py3-none-any.whl (8.8 kB view details)

Uploaded Python 3

File details

Details for the file harness_ccm_external_data-0.1.5.tar.gz.

File metadata

  • Download URL: harness_ccm_external_data-0.1.5.tar.gz
  • Upload date:
  • Size: 7.7 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/2.1.4 CPython/3.13.5 Linux/6.1.134-152.225.amzn2023.x86_64

File hashes

Hashes for harness_ccm_external_data-0.1.5.tar.gz
Algorithm Hash digest
SHA256 328e50efa3613cb0a6b643957f781d4a1188a8ea2ec7f32c9ca07243223eb826
MD5 73d83e87a8b2e336332e263eedcbd195
BLAKE2b-256 fb4975a79e482b33ff957ae4fc429a4be0c935688c83e7937b8e786d2062c7a9

See more details on using hashes here.

File details

Details for the file harness_ccm_external_data-0.1.5-py3-none-any.whl.

File metadata

File hashes

Hashes for harness_ccm_external_data-0.1.5-py3-none-any.whl
Algorithm Hash digest
SHA256 b46e2994ec677aa5da3a5c68b418a51df836e5a66902c9ccdd37594918257eeb
MD5 3f5cbd296aa4f02e463ffb23aad07586
BLAKE2b-256 7458857d59bd667ec557ffac21acec2b90d06cb23c71a92950ef6c8eff43ebcb

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page