Decouple Torch Network-Aware Training on Interlinked Online Nodes (DeToNATION)

These details have not been verified by PyPI

Project links

License
- OSI Approved :: MIT License
Operating System
- OS Independent
Programming Language
- Python :: 3

Project description

Decoupled Torch Network-Aware Training on Interlinked Online Nodes (DeToNATION)

Installation

Installation from PyPI:

pip install detonation

Installation from source:

git clone https://github.com/schneiderkamplab/DeToNATION
cd DeToNATION
pip install .

Usage

The usage requires three elements as exemplified below for using the FlexDeMo optimizer.

First, you need to wrap your model with FSDP and the hybrid sharding strategy:

from torch.distributed.fsdp import FullyShardedDataParallel as FSDP
model = FSDP(
    model,
    sharding_strategy=ShardingStrategy.HYBRID_SHARD,
)

Then, you can import and instantiate the FlexDeMo optimizer:

from detonation import DeMo
optim = DeMo(
    compression_topk=16,
    compression_chunk=128,
    sharding_parallel_group=model.process_group,
    replication_parallel_group=model._inter_node_pg,
)

Third and last, you need to wrap the forward and backward pass using a no_sync context manager to avoid automatic full gradient synchronization:

    with model.no_sync(): # Disable gradient synchronizations across FSDP instances.
        loss = model(input_ids=batch["input_ids"],labels=batch["labels"])["loss"]
        loss.backward()

Project details

These details have not been verified by PyPI

Project links

License
- OSI Approved :: MIT License
Operating System
- OS Independent
Programming Language
- Python :: 3

Release history Release notifications | RSS feed

0.5.2

May 26, 2025

0.5.1

May 23, 2025

0.5.0 yanked

May 23, 2025

Reason this release was yanked:

Discovered bug

0.4.0b0 pre-release

May 1, 2025

0.3.0

Apr 10, 2025

0.2.2

Mar 18, 2025

0.2.1

Feb 12, 2025

This version

0.2.0

Feb 12, 2025

0.1.1

Feb 11, 2025

0.1.0

Feb 10, 2025

0.0.2

Feb 6, 2025

0.0.1

Feb 6, 2025

0.0.0

Jan 23, 2025

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

detonation-0.2.0.tar.gz (7.7 kB view details)

Uploaded Feb 12, 2025 Source

File details

Details for the file detonation-0.2.0.tar.gz.

File metadata

Download URL: detonation-0.2.0.tar.gz
Upload date: Feb 12, 2025
Size: 7.7 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.12.9

File hashes

Hashes for detonation-0.2.0.tar.gz
Algorithm	Hash digest
SHA256	`0264cb221bb5fba5ff69d8fc67c287105aac2e3e35b15da42483bdcc70ef3ac4`
MD5	`d8fc49358f7ba907d66a51d8da775c30`
BLAKE2b-256	`b5056928bad75bcd14b426fe4859f5298353e4571b2b7fd33c0925751d9c3be4`

See more details on using hashes here.

detonation 0.2.0

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Decoupled Torch Network-Aware Training on Interlinked Online Nodes (DeToNATION)

Installation

Usage

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

File details

File metadata

File hashes