Skip to main content

FhirSheets is a command-line tool that reads an Excel file in FHIR cohort format and generates FHIR bundle JSON files from it. Each row in the template Excel file is used to create an individual JSON file, outputting them to a specified folder.

Project description

FHIRSheets

FhirSheets is a command-line tool that reads an Excel file in FHIR cohort format and generates FHIR bundle JSON files from it. Each row in the template Excel file is used to create an individual JSON file, outputting them to a specified folder.

Table of Contents

Features

  • Reads an Excel file following the FHIR cohort import template.
  • Converts each row in the Excel file to a FHIR bundle JSON file.
  • Exports generated JSON files to a specified output folder.

Requirements

  • Python 3.x
  • Required Python packages (see requirements.txt)

Installation

  1. Clone this repository:
    git clone https://github.com/CDCgov/synthetic-data.git
    cd fhir-python-cohort-generation
    
  2. Install the required packages:
    pip install -r requirements.txt
    
    Or use poetry
    poetry build
    

Usage

  1. Fill Out the Template:

    • Open the template file src/resources/Fhir_Cohort_Import_Template.xlsx.
    • Fill out each row with the relevant data.
  2. Run the Tool:

    • Use the python -m src.cli.fhirsheets module script with the required arguments:
      • --input-file: The path to the input Excel file.
      • --output-folder: The path to the output folder where the JSON files will be saved.
    python -m src.fhir_sheets.cli.main --input_file src/resources/Fhir_Cohort_Import_Template.xlsx --output_folder /path/to/output/folder
    
  3. The tool will generate one FHIR bundle JSON file for each row defined in the template.

Example

python -m src.fhir_sheets.cli.main --input_file src/resources/Fhir_Cohort_Import_Template.xlsx --output_folder ./output_bundles

In this example, each row in the Fhir_Cohort_Import_Template.xlsx file will be processed, and a corresponding JSON file will be generated in the output_bundles folder.

Configuration

FHIRSheets supports configuration through JSON files or command-line arguments to customize behavior, including default resource references.

Using a Configuration File

Create a JSON configuration file (e.g., my_config.json):

{
  "enable_default_resource_links": true,
  "default_resource_references": [
    ["observation", "patient", "subject"],
    ["procedure", "patient", "subject"]
  ]
}

Then use it with the CLI:

python -m src.fhir_sheets.cli.main --input_file input.xlsx --output_folder output/ --config_file my_config.json

Configuration Options

  • enable_default_resource_links (boolean, default: true): Enable/disable automatic default resource linking
  • default_resource_references (array): Customize which default resource references to create automatically
  • array_type_references (array): Specify which references should be arrays
  • preview_mode (boolean, default: false): Generate resources in preview mode
  • medications_as_reference (boolean, default: false): Convert medicationCodeableConcept to medication resources
  • build_empty_resources (boolean, default: false): Build resources even when no data exists

Command-Line Arguments

You can also pass configuration options directly:

python -m src.fhir_sheets.cli.main \
  --input_file input.xlsx \
  --output_folder output/ \
  --enable_default_resource_links true \
  --build_empty_resources false

Or use the convenient flag to disable default resource links:

python -m src.fhir_sheets.cli.main \
  --input_file input.xlsx \
  --output_folder output/ \
  --no-default-links

Note: For advanced configuration like customizing default_resource_references lists, use a JSON config file.

Example Configuration File

See config_example.json for a complete example with all default values and available options.

Programmatic Usage

When using FHIRSheets as a Python library, use the simplified context-based API:

from fhir_sheets.core.config.FhirSheetsConfiguration import FhirSheetsConfiguration
from fhir_sheets.core.conversion import create_transaction_bundle, ConversionContext
from fhir_sheets.core import read_input
import json

# Load configuration from JSON file
with open('my_config.json', 'r') as f:
    config_dict = json.load(f)
config = FhirSheetsConfiguration(config_dict)

# Read input data
resource_defs, resource_links, cohort_data = read_input.read_xlsx_and_process("input.xlsx")

# Create context object (encapsulates all parameters)
ctx = ConversionContext(
    resource_definitions=resource_defs,
    resource_links=resource_links,
    cohort_data=cohort_data,
    index=0,
    config=config
)

# Generate bundle with simplified signature
bundle = create_transaction_bundle(ctx)

Processing Multiple Patients

The context-based approach makes it easy to process multiple patients:

from fhir_sheets.core.conversion import create_transaction_bundle, ConversionContext

# Read input data once
resource_defs, resource_links, cohort_data = read_input.read_xlsx_and_process("input.xlsx")

# Process all patients
bundles = []
for patient_index in range(len(cohort_data.patients)):
    ctx = ConversionContext(
        resource_definitions=resource_defs,
        resource_links=resource_links,
        cohort_data=cohort_data,
        index=patient_index,
        config=config
    )
    bundle = create_transaction_bundle(ctx)
    bundles.append(bundle)

License

This project is licensed under the MIT License. See the LICENSE file for more information.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

fhir_sheets-2.6.1.tar.gz (27.7 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

fhir_sheets-2.6.1-py3-none-any.whl (34.9 kB view details)

Uploaded Python 3

File details

Details for the file fhir_sheets-2.6.1.tar.gz.

File metadata

  • Download URL: fhir_sheets-2.6.1.tar.gz
  • Upload date:
  • Size: 27.7 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.13.7

File hashes

Hashes for fhir_sheets-2.6.1.tar.gz
Algorithm Hash digest
SHA256 6bbbdbe1a867a046dfaf4812209f6849eb5983c29c3dca4259bc856990868982
MD5 fe5930a17fa656e37cd37e5f1b3dcccd
BLAKE2b-256 45ecd710dad30e6d40bcdc1be24b33fb4034be03f7b23ba937584abe7b8ff03a

See more details on using hashes here.

File details

Details for the file fhir_sheets-2.6.1-py3-none-any.whl.

File metadata

  • Download URL: fhir_sheets-2.6.1-py3-none-any.whl
  • Upload date:
  • Size: 34.9 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.13.7

File hashes

Hashes for fhir_sheets-2.6.1-py3-none-any.whl
Algorithm Hash digest
SHA256 1b3766933b8864d57ec5ded517d156584bb371ff43ad0d32892531a8fb3424cd
MD5 a5088af43124b087107ee1bff3645303
BLAKE2b-256 52a8e7843995f4822343ddec6c4947c07f81c55e655aa021718d4e1e5e85e8e5

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page