torchmetrics

PyTorch native Metrics

These details have not been verified by PyPI

Project links

Project description

Machine learning metrics for distributed, scalable PyTorch applications.

What is Torchmetrics • Implementing a metric • Built-in metrics • Docs • Community • License

Installation

Simple installation from PyPI

pip install torchmetrics

Other installations

Install using conda

conda install -c conda-forge torchmetrics

Install using uv

uv add torchmetrics

Pip from source

# with git
pip install git+https://github.com/Lightning-AI/torchmetrics.git@release/stable

Pip from archive

pip install https://github.com/Lightning-AI/torchmetrics/archive/refs/heads/release/stable.zip

Extra dependencies for specialized metrics:

pip install torchmetrics[audio]
pip install torchmetrics[image]
pip install torchmetrics[text]
pip install torchmetrics[all]  # install all of the above

Install latest developer version

pip install https://github.com/Lightning-AI/torchmetrics/archive/master.zip

What is TorchMetrics

TorchMetrics is a collection of 100+ PyTorch metrics implementations and an easy-to-use API to create custom metrics. It offers:

A standardized interface to increase reproducibility
Reduces boilerplate
Automatic accumulation over batches
Metrics optimized for distributed-training
Automatic synchronization between multiple devices

You can use TorchMetrics with any PyTorch model or with PyTorch Lightning to enjoy additional features such as:

Module metrics are automatically placed on the correct device.
Native support for logging metrics in Lightning to reduce even more boilerplate.

Using TorchMetrics

Module metrics

The module-based metrics contain internal metric states (similar to the parameters of the PyTorch module) that automate accumulation and synchronization across devices!

Automatic accumulation over multiple batches
Automatic synchronization between multiple devices
Metric arithmetic

This can be run on CPU, single GPU or multi-GPUs!

For the single GPU/CPU case:

import torch

# import our library
import torchmetrics

# initialize metric
metric = torchmetrics.classification.Accuracy(task="multiclass", num_classes=5)

# move the metric to device you want computations to take place
device = "cuda" if torch.cuda.is_available() else "cpu"
metric.to(device)

n_batches = 10
for i in range(n_batches):
    # simulate a classification problem
    preds = torch.randn(10, 5).softmax(dim=-1).to(device)
    target = torch.randint(5, (10,)).to(device)

    # metric on current batch
    acc = metric(preds, target)
    print(f"Accuracy on batch {i}: {acc}")

# metric on all batches using custom accumulation
acc = metric.compute()
print(f"Accuracy on all data: {acc}")

Module metric usage remains the same when using multiple GPUs or multiple nodes.

Example using DDP

import os
import torch
import torch.distributed as dist
import torch.multiprocessing as mp
from torch import nn
from torch.nn.parallel import DistributedDataParallel as DDP
import torchmetrics


def metric_ddp(rank, world_size):
    os.environ["MASTER_ADDR"] = "localhost"
    os.environ["MASTER_PORT"] = "12355"

    # create default process group
    dist.init_process_group("gloo", rank=rank, world_size=world_size)

    # initialize model
    metric = torchmetrics.classification.Accuracy(task="multiclass", num_classes=5)

    # define a model and append your metric to it
    # this allows metric states to be placed on correct accelerators when
    # .to(device) is called on the model
    model = nn.Linear(10, 10)
    model.metric = metric
    model = model.to(rank)

    # initialize DDP
    model = DDP(model, device_ids=[rank])

    n_epochs = 5
    # this shows iteration over multiple training epochs
    for n in range(n_epochs):
        # this will be replaced by a DataLoader with a DistributedSampler
        n_batches = 10
        for i in range(n_batches):
            # simulate a classification problem
            preds = torch.randn(10, 5).softmax(dim=-1)
            target = torch.randint(5, (10,))

            # metric on current batch
            acc = metric(preds, target)
            if rank == 0:  # print only for rank 0
                print(f"Accuracy on batch {i}: {acc}")

        # metric on all batches and all accelerators using custom accumulation
        # accuracy is same across both accelerators
        acc = metric.compute()
        print(f"Accuracy on all data: {acc}, accelerator rank: {rank}")

        # Resetting internal state such that metric ready for new data
        metric.reset()

    # cleanup
    dist.destroy_process_group()


if __name__ == "__main__":
    world_size = 2  # number of gpus to parallelize over
    mp.spawn(metric_ddp, args=(world_size,), nprocs=world_size, join=True)

Implementing your own Module metric

Implementing your own metric is as easy as subclassing an torch.nn.Module. Simply, subclass torchmetrics.Metric and just implement the update and compute methods:

import torch
from torchmetrics import Metric


class MyAccuracy(Metric):
    def __init__(self):
        # remember to call super
        super().__init__()
        # call `self.add_state`for every internal state that is needed for the metrics computations
        # dist_reduce_fx indicates the function that should be used to reduce
        # state from multiple processes
        self.add_state("correct", default=torch.tensor(0), dist_reduce_fx="sum")
        self.add_state("total", default=torch.tensor(0), dist_reduce_fx="sum")

    def update(self, preds: torch.Tensor, target: torch.Tensor) -> None:
        # extract predicted class index for computing accuracy
        preds = preds.argmax(dim=-1)
        assert preds.shape == target.shape
        # update metric states
        self.correct += torch.sum(preds == target)
        self.total += target.numel()

    def compute(self) -> torch.Tensor:
        # compute final result
        return self.correct.float() / self.total


my_metric = MyAccuracy()
preds = torch.randn(10, 5).softmax(dim=-1)
target = torch.randint(5, (10,))

print(my_metric(preds, target))

Functional metrics

Similar to torch.nn, most metrics have both a module-based and functional version. The functional versions are simple python functions that as input take torch.tensors and return the corresponding metric as a torch.tensor.

import torch

# import our library
import torchmetrics

# simulate a classification problem
preds = torch.randn(10, 5).softmax(dim=-1)
target = torch.randint(5, (10,))

acc = torchmetrics.functional.classification.multiclass_accuracy(
    preds, target, num_classes=5
)

Covered domains and example metrics

In total TorchMetrics contains 100+ metrics, which covers the following domains:

Audio
Classification
Detection
Information Retrieval
Image
Multimodal (Image-Text-3D Talking Heads)
Nominal
Regression
Segmentation
Text

Each domain may require some additional dependencies which can be installed with pip install torchmetrics[audio], pip install torchmetrics['image'] etc.

Additional features

Plotting

Visualization of metrics can be important to help understand what is going on with your machine learning algorithms. Torchmetrics have built-in plotting support (install dependencies with pip install torchmetrics[visual]) for nearly all modular metrics through the .plot method. Simply call the method to get a simple visualization of any metric!

import torch
from torchmetrics.classification import MulticlassAccuracy, MulticlassConfusionMatrix

num_classes = 3

# this will generate two distributions that comes more similar as iterations increase
w = torch.randn(num_classes)
target = lambda it: torch.multinomial((it * w).softmax(dim=-1), 100, replacement=True)
preds = lambda it: torch.multinomial((it * w).softmax(dim=-1), 100, replacement=True)

acc = MulticlassAccuracy(num_classes=num_classes, average="micro")
acc_per_class = MulticlassAccuracy(num_classes=num_classes, average=None)
confmat = MulticlassConfusionMatrix(num_classes=num_classes)

# plot single value
for i in range(5):
    acc_per_class.update(preds(i), target(i))
    confmat.update(preds(i), target(i))
fig1, ax1 = acc_per_class.plot()
fig2, ax2 = confmat.plot()

# plot multiple values
values = []
for i in range(10):
    values.append(acc(preds(i), target(i)))
fig3, ax3 = acc.plot(values)

For examples of plotting different metrics try running this example file.

Contribute!

The lightning + TorchMetrics team is hard at work adding even more metrics. But we're looking for incredible contributors like you to submit new metrics and improve existing ones!

Join our Discord to get help with becoming a contributor!

Community

For help or questions, join our huge community on Discord!

Citation

We’re excited to continue the strong legacy of open source software and have been inspired over the years by Caffe, Theano, Keras, PyTorch, torchbearer, ignite, sklearn and fast.ai.

If you want to cite this framework feel free to use GitHub's built-in citation option to generate a bibtex or APA-Style citation based on this file (but only if you loved it 😊).

License

Please observe the Apache 2.0 license that is listed in this repository. In addition, the Lightning framework is Patent Pending.

Project details

These details have not been verified by PyPI

Project links

Release history Release notifications | RSS feed

This version

1.8.2

Sep 3, 2025

1.8.1

Aug 7, 2025

1.8.0

Jul 23, 2025

1.7.4

Jul 5, 2025

1.7.3

Jun 13, 2025

1.7.2

May 28, 2025

1.7.1

Apr 7, 2025

1.7.0

Mar 20, 2025

1.6.3

Mar 14, 2025

1.6.2

Mar 3, 2025

1.6.1

Dec 25, 2024

1.6.0

Nov 12, 2024

1.5.2

Nov 8, 2024

1.5.1

Oct 23, 2024

1.5.0

Oct 18, 2024

1.4.3

Oct 10, 2024

1.4.2

Sep 13, 2024

1.4.1

Aug 3, 2024

1.4.0.post0

May 15, 2024

1.4.0

May 6, 2024

1.3.2

Mar 18, 2024

1.3.1

Feb 12, 2024

1.3.0.post0

Jan 17, 2024

1.3.0 yanked

Jan 11, 2024

1.2.1

Dec 1, 2023

1.2.0

Sep 22, 2023

1.1.2

Sep 11, 2023

1.1.1

Aug 29, 2023

1.1.0

Aug 22, 2023

1.0.3

Aug 8, 2023

1.0.2

Aug 3, 2023

1.0.1

Jul 13, 2023

1.0.0

Jul 5, 2023

1.0.0rc1 pre-release

Jun 29, 2023

1.0.0rc0 pre-release

May 4, 2023

0.11.4

Mar 10, 2023

0.11.3

Feb 28, 2023

0.11.2

Feb 28, 2023

0.11.1

Jan 31, 2023

0.11.0

Nov 30, 2022

0.10.3

Nov 16, 2022

0.10.2

Oct 31, 2022

0.10.1

Oct 26, 2022

0.10.0

Oct 4, 2022

0.10.0rc0 pre-release

Sep 19, 2022

0.9.3

Jul 23, 2022

0.9.2

Jun 29, 2022

0.9.1

Jun 8, 2022

0.9.0

May 31, 2022

0.8.2

May 6, 2022

0.8.1

Apr 27, 2022

0.8.0

Apr 15, 2022

0.8.0rc0 pre-release

Apr 13, 2022

0.7.3

Mar 23, 2022

0.7.2

Feb 10, 2022

0.7.1

Feb 3, 2022

0.7.0

Jan 17, 2022

0.7.0rc1 pre-release

Jan 15, 2022

0.7.0rc0 pre-release

Jan 13, 2022

0.6.2

Dec 15, 2021

0.6.1

Dec 6, 2021

0.6.0

Oct 28, 2021

0.6.0rc1 pre-release

Oct 28, 2021

0.6.0rc0 pre-release

Oct 25, 2021

0.5.1

Sep 1, 2021

0.5.0

Aug 10, 2021

0.4.1

Jul 5, 2021

0.4.0 yanked

Jun 29, 2021

Reason this release was yanked:

DDP

0.3.2

May 10, 2021

0.3.1

Apr 21, 2021

0.3.0

Apr 20, 2021

0.2.0

Mar 12, 2021

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

torchmetrics-1.8.2.tar.gz (580.7 kB view details)

Uploaded Sep 3, 2025 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

torchmetrics-1.8.2-py3-none-any.whl (983.2 kB view details)

Uploaded Sep 3, 2025 Python 3

File details

Details for the file torchmetrics-1.8.2.tar.gz.

File metadata

Download URL: torchmetrics-1.8.2.tar.gz
Upload date: Sep 3, 2025
Size: 580.7 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.12.8

File hashes

Hashes for torchmetrics-1.8.2.tar.gz
Algorithm	Hash digest
SHA256	`cf64a901036bf107f17a524009eea7781c9c5315d130713aeca5747a686fe7a5`
MD5	`7eae61deccf78c5821bf68ac175e3f44`
BLAKE2b-256	`852e48a887a59ecc4a10ce9e8b35b3e3c5cef29d902c4eac143378526e7485cb`

See more details on using hashes here.

File details

Details for the file torchmetrics-1.8.2-py3-none-any.whl.

File metadata

Download URL: torchmetrics-1.8.2-py3-none-any.whl
Upload date: Sep 3, 2025
Size: 983.2 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.12.8

File hashes

Hashes for torchmetrics-1.8.2-py3-none-any.whl
Algorithm	Hash digest
SHA256	`08382fd96b923e39e904c4d570f3d49e2cc71ccabd2a94e0f895d1f0dac86242`
MD5	`c18d0de8b65d8e4b8f2f98c59debb667`
BLAKE2b-256	`0221aa0f434434c48490f91b65962b1ce863fdcce63febc166ca9fe9d706c2b6`

See more details on using hashes here.

torchmetrics 1.8.2

Navigation

Verified details

Owner

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Installation

What is TorchMetrics

Using TorchMetrics

Module metrics

Implementing your own Module metric

Functional metrics

Covered domains and example metrics

Additional features

Plotting

Contribute!

Community

Citation

License

Project details

Verified details

Owner

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes