Skip to main content

A Python package for uploading training metrics and checkpointing data to Pico backend databases

Project description

Pico Report

A Python package for uploading training metrics and checkpointing data to Pico backend databases. This package allows you to seamlessly integrate with your existing training pipelines while sending data to your private Pico dashboard.

Installation

pip install pico-report

Quick Start

First, set up your environment variables (recommended) or use direct configuration.

Setup Environment Variables

# Copy the example file
cp .env.example .env

# Edit .env with your actual credentials (NEVER commit this file!)
# PICO_API_KEY=your_actual_api_key
# PICO_LAB_HASH=your_actual_lab_hash

Using the Client

from pico_report import PicoClient, ReporterConfig

# Method 1: Using environment variables (recommended - secure)
client = PicoClient()

# Method 2: Direct configuration (not recommended for production)
config = ReporterConfig(
    api_key="your-api-key",  # Required
    lab_hash="your-lab-hash",  # Required
    experiment_name="experiment-1"  # Optional
)
client = PicoClient(config=config)

# Log training metrics
client.log_metrics({
    "loss": 0.5,
    "accuracy": 0.85,
    "learning_rate": 0.001
}, step=100)

# Upload checkpoint data
client.upload_checkpoint_data({
    "model_state": "path/to/checkpoint",
    "optimizer_state": "path/to/optimizer",
    "epoch": 5
}, step=100)

Configuration

Environment Variables

Set the following required environment variables:

export PICO_API_KEY="your-api-key"
export PICO_LAB_HASH="your-lab-hash"

Optional configuration:

# Automatically create git commits for each experiment (default: false)
export PICO_AUTO_COMMIT="true"

# Experiment name (if not provided programmatically)
export PICO_EXPERIMENT_NAME="my-experiment"

Configuration File

Create a .env file in your project root:

# Required
PICO_API_KEY=your-api-key
PICO_LAB_HASH=your-lab-hash

# Optional
PICO_EXPERIMENT_NAME=my-experiment
PICO_AUTO_COMMIT=true

Git Integration & Auto-Commit

Pico Report can automatically create Git commits when you create experiments, allowing you to track the exact code state used for each run.

Enabling Auto-Commit

from pico_report.integrations import PicoReporter

# Method 1: Via environment variable
# Set PICO_AUTO_COMMIT=true in your .env file

# Method 2: Via configuration
reporter = PicoReporter(
    lab_hash="my-lab-hash",
    auto_commit=True  # Enable automatic git commits
)

# Setup experiment - will auto-commit if enabled
reporter.setup_experiment(
    experiment_name="my-experiment",
    config_data={"lr": 0.001}
)
# Git commit created automatically with message: "Experiment: my-experiment"
# Commit is pushed to your remote repository (if configured)

Requirements for Auto-Commit

  • Your code must be in a Git repository
  • Git must be installed and available in PATH
  • For automatic push: Git remote must be configured with authentication

What Gets Committed

When auto-commit is enabled:

  1. All modified and new files are staged (git add -A)
  2. A commit is created with message: "Experiment: {experiment_name}"
  3. The commit SHA is linked to your experiment in the dashboard
  4. The commit is automatically pushed to your remote repository
  5. You can view code diffs between experiments in the Pico Labs UI

Disabling Auto-Commit

# Disable for specific reporter
reporter = PicoReporter(
    lab_hash="my-lab-hash",
    auto_commit=False  # Disable automatic commits
)

# Or set environment variable
# PICO_AUTO_COMMIT=false

High-Level Interface

For easier integration, use the PicoReporter class:

from pico_report.integrations import PicoReporter

# lab_hash is required - provide it explicitly or via PICO_LAB_HASH environment variable
reporter = PicoReporter(
    lab_hash="my-lab-hash",  # Required
    experiment_name="transformer-training",  # Optional
    auto_commit=True  # Optional: enable automatic git commits (default: False)
)

# Setup experiment (auto-commits if enabled)
reporter.setup_experiment(
    experiment_name="transformer-training",
    config_data={"lr": 0.001, "batch_size": 32},
    description="Training transformer model"
)
# If auto_commit=True, a git commit is automatically created and pushed

# Log training metrics
reporter.log_training_metrics({
    "loss": 0.5,
    "perplexity": 2.1
}, step=100)

# Log evaluation metrics
reporter.log_evaluation_metrics({
    "eval_loss": 0.45,
    "eval_accuracy": 0.87
}, step=100, prefix="validation")

Integration with Existing Training Code

PyTorch Lightning Integration

import lightning as L
from pico_report.integrations import PicoReporter

class MyLightningModule(L.LightningModule):
    def __init__(self):
        super().__init__()
        # Requires PICO_API_KEY and PICO_LAB_HASH environment variables to be set
        self.pico_reporter = PicoReporter(
            experiment_name="lightning-training",
            auto_commit=True  # Track code changes automatically
        )
    
    def training_step(self, batch, batch_idx):
        # Your training logic
        loss = self.compute_loss(batch)
        
        # Log to Pico
        if self.global_step % 10 == 0:
            self.pico_reporter.log_training_metrics({
                "train_loss": loss.item()
            }, step=self.global_step)
        
        return loss
    
    def validation_step(self, batch, batch_idx):
        # Your validation logic
        val_loss = self.compute_loss(batch)
        return {"val_loss": val_loss}
    
    def validation_epoch_end(self, outputs):
        avg_loss = torch.stack([x["val_loss"] for x in outputs]).mean()
        
        self.pico_reporter.log_evaluation_metrics({
            "val_loss": avg_loss.item()
        }, step=self.global_step)

Direct Integration with Pico-Train

You can modify your existing pico-train setup to also send data to Pico backend:

# In your training script
from pico_report.integrations import PicoReporter

# Initialize both wandb and pico reporter
wandb_logger = initialize_wandb(monitoring_config, checkpointing_config)
pico_reporter = PicoReporter(
    lab_hash=monitoring_config.pico.lab_hash,
    experiment_name=checkpointing_config.run_name,
    auto_commit=True  # Automatically commit experiment code
)

# In your training loop
for step, batch in enumerate(dataloader):
    # ... training logic ...
    
    # Log to both wandb and pico
    metrics = {"loss": loss.item(), "lr": lr}
    wandb_logger.log(metrics, step=step)
    pico_reporter.log_training_metrics(metrics, step=step)

Configuration

The base URL should point to the report API endpoint. For local development:

export PICO_BASE_URL="http://localhost:3000/api/report"

For production:

export PICO_BASE_URL="https://picolabs.space/api/report"

API Reference

PicoClient

Main client for direct API interaction.

Methods

  • log_metrics(metrics, step, timestamp): Log training metrics
  • create_experiment(name, config_data, description): Create new experiment
  • list_experiments(limit, offset): List existing experiments

PicoReporter

High-level interface for easier integration.

Methods

  • setup_experiment(name, config_data, description): Setup experiment
  • log_training_metrics(metrics, step, prefix): Log training metrics with prefix
  • log_evaluation_metrics(metrics, step, prefix): Log evaluation metrics with prefix
  • log_analysis_metrics(metric_name, metric_data, step, data_split, prefix): Log learning dynamics analysis metrics
  • log_system_metrics(**metrics): Log system performance metrics

Error Handling

The package includes custom exceptions:

  • PicoReportError: Base exception
  • PicoAuthError: Authentication related errors
  • PicoUploadError: Data upload errors
  • PicoConfigError: Configuration errors
  • PicoGitError: Git operations errors (when using auto-commit)
from pico_report.exceptions import PicoAuthError, PicoUploadError, PicoGitError

try:
    client.log_metrics(metrics, step=100)
except PicoAuthError:
    print("Authentication failed - check your API key")
except PicoUploadError as e:
    print(f"Upload failed: {e}")
except PicoGitError as e:
    print(f"Git operation failed: {e}")
    print("Note: Experiment was created, but git commit failed")

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

pico_report-1.1.1.tar.gz (14.0 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

pico_report-1.1.1-py3-none-any.whl (15.2 kB view details)

Uploaded Python 3

File details

Details for the file pico_report-1.1.1.tar.gz.

File metadata

  • Download URL: pico_report-1.1.1.tar.gz
  • Upload date:
  • Size: 14.0 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/2.1.3 CPython/3.13.3 Darwin/24.6.0

File hashes

Hashes for pico_report-1.1.1.tar.gz
Algorithm Hash digest
SHA256 ebe0bcf1e6b659473b0b42ba8ae0a16a819a7fdff7734c344f8b8778d245117f
MD5 e193151f7cae2afc5274410361ca8448
BLAKE2b-256 003f759e469b08efa7c36d866fab2f5c3e0e93422af33ad7fada11aeb8ce2439

See more details on using hashes here.

File details

Details for the file pico_report-1.1.1-py3-none-any.whl.

File metadata

  • Download URL: pico_report-1.1.1-py3-none-any.whl
  • Upload date:
  • Size: 15.2 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/2.1.3 CPython/3.13.3 Darwin/24.6.0

File hashes

Hashes for pico_report-1.1.1-py3-none-any.whl
Algorithm Hash digest
SHA256 e45ae54920ef9dbf50402e8b8fca58c782ebad3d4f5bb4e415f68df4c041d6a1
MD5 eae60f1884debe57e9383e01a653c953
BLAKE2b-256 c80efc2ad4e67321107076c8bf841c6753702499b44d2bfe3ff87637f3309038

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page