Skip to main content

HydraFlow seamlessly integrates Hydra and MLflow to streamline ML experiment management, combining Hydra's configuration management with MLflow's tracking capabilities.

Project description

HydraFlow

PyPI Version Build Status Coverage Status Documentation Status Python Version

Overview

HydraFlow seamlessly integrates Hydra and MLflow to streamline machine learning experiment workflows. By combining Hydra's powerful configuration management with MLflow's robust experiment tracking, HydraFlow provides a comprehensive solution for defining, executing, and analyzing machine learning experiments.

Design Principles

HydraFlow is built on the following design principles:

  1. Type Safety - Utilizing Python dataclasses for configuration type checking and IDE support
  2. Reproducibility - Automatically tracking all experiment configurations for fully reproducible experiments
  3. Analysis Capabilities - Providing powerful APIs for easily analyzing experiment results
  4. Workflow Integration - Creating a cohesive workflow by integrating Hydra's configuration management with MLflow's experiment tracking

Key Features

  • Type-safe Configuration Management - Define experiment parameters using Python dataclasses with full IDE support and validation
  • Seamless Hydra-MLflow Integration - Automatically register configurations with Hydra and track experiments with MLflow
  • Advanced Parameter Sweeps - Define complex parameter spaces using extended sweep syntax for numerical ranges, combinations, and SI prefixes
  • Workflow Automation - Create reusable experiment workflows with YAML-based job definitions
  • Powerful Analysis Tools - Filter, group, and analyze experiment results with type-aware APIs
  • Custom Implementation Support - Extend experiment analysis with domain-specific functionality

Installation

pip install hydraflow

Requirements: Python 3.13+

Quick Example

import hydraflow
from dataclasses import dataclass
from mlflow.entities import Run

@dataclass
class Config:
    width: int = 1024
    height: int = 768

@hydraflow.main(Config, tracking_uri="sqlite:///mlflow.db")
def app(run: Run, cfg: Config) -> None:
    # Your experiment code here
    print(f"Running with width={cfg.width}, height={cfg.height}")

if __name__ == "__main__":
    app()

Execute a parameter sweep with:

python app.py -m width=800,1200 height=600,900

Core Components

HydraFlow consists of the following key components:

Configuration Management

Define type-safe configurations using Python dataclasses:

@dataclass
class Config:
    learning_rate: float = 0.001
    batch_size: int = 32
    epochs: int = 10

Main Decorator

The @hydraflow.main decorator integrates Hydra and MLflow:

@hydraflow.main(Config)
def train(run: Run, cfg: Config) -> None:
    # Your experiment code

Workflow Automation

Define reusable experiment workflows in YAML:

jobs:
  train_models:
    run: python train.py
    sets:
      - each: model=small,medium,large
        all: learning_rate=0.001,0.01,0.1

Analysis Tools

Analyze experiment results with powerful APIs:

import mlflow
from hydraflow import Run, iter_run_dirs

# Load runs
mlflow.set_tracking_uri("sqlite:///mlflow.db")
runs = Run.load(iter_run_dirs())

# Filter and analyze
best_runs = runs.filter(model_type="transformer").to_frame("learning_rate", "accuracy")

Documentation

For detailed documentation, visit our documentation site:

License

This project is licensed under the MIT License.

Project details


Release history Release notifications | RSS feed

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

hydraflow-0.21.2.tar.gz (31.0 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

hydraflow-0.21.2-py3-none-any.whl (38.9 kB view details)

Uploaded Python 3

File details

Details for the file hydraflow-0.21.2.tar.gz.

File metadata

  • Download URL: hydraflow-0.21.2.tar.gz
  • Upload date:
  • Size: 31.0 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: uv/0.9.13 {"installer":{"name":"uv","version":"0.9.13"},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"Ubuntu","version":"24.04","id":"noble","libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":true}

File hashes

Hashes for hydraflow-0.21.2.tar.gz
Algorithm Hash digest
SHA256 69a24a2202b86cc1a4ed744fc9225977561ffab617a84b52ac0d63283184ed5e
MD5 0ac4ec80aa6b3628872cc38014303ce5
BLAKE2b-256 35ac4e241ec1934a81ad6ae413328d4f2838c85d62d02a4768e2bb3bbf612a80

See more details on using hashes here.

File details

Details for the file hydraflow-0.21.2-py3-none-any.whl.

File metadata

  • Download URL: hydraflow-0.21.2-py3-none-any.whl
  • Upload date:
  • Size: 38.9 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: uv/0.9.13 {"installer":{"name":"uv","version":"0.9.13"},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"Ubuntu","version":"24.04","id":"noble","libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":true}

File hashes

Hashes for hydraflow-0.21.2-py3-none-any.whl
Algorithm Hash digest
SHA256 7a10874deef0c41ab1cb6faf81ee5419e87c1f73bfd0f2c581fedcb8e74188ec
MD5 03ef57cb26b9744d369a1ad05a4f54a1
BLAKE2b-256 58e8d53b1bbdf0848a095c8ccac02d5f71f37c4069bb4e6308e355731cb48799

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page