Skip to main content

Generic Bayesian self-optimizing hyperparameter tuning engine

Project description

image

AutoTuneNet

A generic, open-source Python library that enables self-optimizing model training by dynamically tuning hyperparameters during training using Bayesian Optimization.

Instead of traditional manually tuning learning rates, batch sizes, or regularization values before training, AutoTuneNet continuously observes training behavior and automatically adjusts hyperparameters to improve convergence and performance. AutoTuneNet combines Bayesian optimization with safety-aware adaptation to tune hyperparameters during training without destabilizing optimization. It is designed for scenarios where stability and restart free training matter.

Why This Project Exists

Hyperparameter tuning is one of the most time-consuming and error-prone parts of machine learning workflows.

Common problems:

  • Manual trial-and-error
  • Grid/random search waste compute
  • Hyperparameters are fixed before training
  • Optimal values often change during training

AutoTuneNet solves this by making hyperparameter tuning part of the training loop itself.

Core Idea

AutoTuneNet treats hyperparameter tuning as a learning problem.

During training:

  1. The model trains normally
  2. Training and validation metrics are observed
  3. A Bayesian optimizer models the relationship between hyperparameters and performance
  4. Hyperparameters are updated incrementally and safely
  5. Training continues with improved settings

This creates a closed-loop, self-optimizing training system.

What This Is (and Is Not)

This project is

  • A generic hyperparameter optimization engine
  • Model-agnostic
  • Dataset-agnostic
  • Designed to plug into existing training loops
  • Suitable for research and production workflows

This project is not

  • A single ML model
  • Offline AutoML that runs many full trials
  • Grid or random search
  • Neural Architecture Search

Design Philosophy

  • Framework-agnostic core
    The Bayesian optimization logic does not depend on PyTorch or TensorFlow.

  • Thin framework adapters
    Framework-specific code lives in adapters (PyTorch first).

  • Safety first
    Guardrails prevent unstable updates and allow rollback.

  • Minimal user code changes
    Users should be able to integrate this with a few lines of code.

Key features

  • Training time hyperparameter optimization
  • Bayesian Optimization (Optuna-backed, ask-tell)
  • Stability guards with rollback protection
  • Metric Smoothing for noisy signals
  • PyTorch Adapter
  • Multi-parameter tuning(lr, momentum, weight_decay etc.)
  • Config-driven tuning via YAML or dict
  • Fully Unit Tested
  • Lightweight & Modular

Installation

pip install autotunenet

Quick Usage

import torch 
import torch.nn as nn
import torch.optim as optim

from autotunenet.core.parameters import ParameterSpace
from autotunenet.core.bayesian_optimizer import BayesianOptimizer
from autotunenet.adapters.pytorch.adapter import PyTorchHyperparameterAdapter

model = nn.Linear(10, 1)
optimizer = optim.Adam(model.parameters(), lr=0.01)

param_space = ParameterSpace({
    "lr": (1e-4, 1e-1)
})

autotune = BayesianOptimizer(param_space)

adapter = PyTorchHyperparameterAdapter(
    torch_optimizer=optimizer,
    autotune_optimizer=autotune
)

for epoch in range(20):
    train_loss = train_one_epoch(model)
    val_metric = -train_loss  # higher is better

    adapter.on_epoch_end(metric=val_metric)

    print(f"Epoch {epoch} | lr={optimizer.param_groups[0]['lr']:.6f}")

That's it AutoTuneNet will:

  • explore hyperparameters
  • keep the best configuration
  • rollback unsafe updates automatically

How it works?

AutoTuneNet runs a suggest -> observe loop inside training.

  1. Suggest new hyperparameters (Bayesian optimization)
  2. Apply them tentatively
  3. Observe training or validation metric
  4. Accept or rollback based on stability rules

This loop repeats throughout training without breaking it.

Safety and Stability and Support

AutoTuneNet is designed to never destabilize training.

  • Built-in protections:
  • Regression detection
  • Consecutive failure thresholds(patience)
  • Cooldown after rollback
  • Restore last known good configuration
  • Warmup phase
  • Bounded Updates

If a suggested hyperparameter harms training, it is reverted immediately.

It supports

  • PyTorch Adapter or Integration
  • Multi-paramter Tuning
  • Config-Driven Tuning

When does a rollback triggers?

A rollback is triggered when training performance regresses significantly and consistently.

At a high level:

  • Training or validation metrics may be smoothed over a short window
  • The current metric is compared against the best recent value
  • A rollback is triggered if:
    1. The relative regression exceeds a configurable threshold
    2. this condition persists for multiple consecutive tuning steps
  • This prevent rollback on single noisy measurements, and expected short-term fluctuations

What happens during rollback?

When rollback is triggered:

  • The most recetn hyperparameter update is rejected
  • Hyperparameters are restored to the last known good configuration
  • A cooldown period is entered to prevent rapid oscillation
  • So rollback restores tuned hyperparameters not the full optimizer or model state.

Restoring only hyperparameters is a deliberate design choice that balances:

  1. safety
  2. performance
  3. framework independence
  4. restoring optimize state is expensive and framework-specific
  5. most instabilities are caused by unsafe hyperparameters
  6. hyperparameter rollback is sufficient to prevent divergence in practice

Cooldown behavior

After rollback, AutoTuneNet enters a short cooldown window during which:

  • further rollbacks are temporarily suppressed
  • training is allowed to stabilize
  • exploration can resume safely afterward

This avoid repeated rollback loops in noisy regions.

Configuration

All stability parameters are configurable, including:

  1. regression threshold
  2. number of consecutive failures
  3. cooldown length

Evalution and Benchmarks:

The current relase focuses on:

  1. correctness
  2. safety
  3. integration quality
  4. test coverage

Formal benchmarks agains:

  • fixed hyperparameters
  • learning rate schedulers
  • offline hyperparameter search are not yet included. Benchmarking and comparative evaluation are planned, and community contributions in this area are very welcome.

Config-Driven Tuning

AutoTuneNet supports configuration-driven hyperparameter tuning with built-in safety mechanisms. AutoTuneNet can be configured entirely via a tuning config:

tuning: 
  tune_n_steps: 1
  warmup_epochs: 2
  max_delta: 0.5
  • Example
from autotunenet.factory import build_pytorch_autotunenet

adapter = build_pytorch_autotunenet(
  torch_optimizer=torch_optimizer,
  raw_config={
    "parameter_space": {
      "lr": [1e-4, 1e-2]
    },
    "tuning": {
      "tune_n_steps": 2,
      "warmup_epochs": 3,
      "max_delta": 0.5
    }
  }
)

Safety-Aware Adaptation

AutoTuneNet performs online hyperparameter adaptation with built-in safety mechanism:

  • warmup phase: delays tuning until early training stabilizes
  • bounded updates(max_delta): limits how much a hyperparameter can change per step
  • Rollback & Cooldown: prevents repeated destabilizing updates

These mechanisms ensure that adaptive tuning remains bounded and predictable, even under non-stationary training dynamics.

Safety Controls

Parameter Description
warmup_epochs Delays tuning until early training stabilizes
max_delta Limits how much a hyperparameter can change per step
tune_n_steps Controls tuning frequency

Testing

AutoTuneNet is fully unit tested.

python -m pytest -v

Tests cover:

  1. optimizer lifecycle
  2. stability logic
  3. rollback behavior
  4. PyTorch adapter
  5. config loading

Folder Structure

autotunenet/
├── AutoTuneNet/   # Bayesian optimizer, parameter space
├── safeguards/    # Stability and rollback logic
├── adapters/      # Framework integrations (PyTorch)
├── config/        # Config schema & loaders
├── logging/       # Structured logging
├── benchmarks/    # Benchmarks(Fixed_lr, offline_HPO, scheduler, stress_test, autotunenet)

License

MIT License

Contributing

Contributions are welcome.

  • Open issues for bugs or ideas
  • PRs fro improvement or adapters
  • Tests required for new features

Acknowledgements

Built on top of:

  • Optuna
  • PyTorch Inspired by real-world ML systems where stability matters more than speed.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

autotunenet-1.0.2.tar.gz (17.0 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

autotunenet-1.0.2-py3-none-any.whl (18.0 kB view details)

Uploaded Python 3

File details

Details for the file autotunenet-1.0.2.tar.gz.

File metadata

  • Download URL: autotunenet-1.0.2.tar.gz
  • Upload date:
  • Size: 17.0 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.10.1

File hashes

Hashes for autotunenet-1.0.2.tar.gz
Algorithm Hash digest
SHA256 d77921e88fcf8dc410f38074821e13dee17b39c4f3b695338c28aae3343aa8bb
MD5 3885296a75d04b3e957ae8c5d23ef122
BLAKE2b-256 7a8e49a3e3e8c496e607bbaa5dbe66736a1caae3a65831980195ebe15f7e9c09

See more details on using hashes here.

File details

Details for the file autotunenet-1.0.2-py3-none-any.whl.

File metadata

  • Download URL: autotunenet-1.0.2-py3-none-any.whl
  • Upload date:
  • Size: 18.0 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.10.1

File hashes

Hashes for autotunenet-1.0.2-py3-none-any.whl
Algorithm Hash digest
SHA256 2fa20125f64ac0f1d2df7cf694f767e2fc6c9f5247f806f9e09e074f30a2d243
MD5 a460e5767d27e2cf39445f86e1de9352
BLAKE2b-256 5d06dd46ed1720622e586502b7d687d2de5f9965c09d2739ac9ff906768afd4a

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page