Location-aware tensor Mahalanobis anomaly detection
Project description
tensor-md
tensor-md detects and localizes unusual image regions from normal training
images. It keeps each local observation as a tensor and scores it with a
location-aware tensor Mahalanobis model using separable covariance factors.
The package is label-free during training: provide a folder of normal images, fit the detector, and then score images from another folder. Larger scores mean that a patch differs more strongly from the normal variation learned at the same spatial location.
Installation
Install the core package from PyPI:
python -m pip install tensor-md
The PyPI distribution is named tensor-md; import it in Python as
tensor_md, because Python module names cannot contain hyphens.
Install the optional CNN dependencies when using the built-in feature extractors:
python -m pip install "tensor-md[cnn]"
Diagnostics and notebooks are available through the evaluation and
notebooks extras.
Quick start
The following example learns directly from image patches. The training and scoring directories may contain PNG, JPEG, BMP, or TIFF images; no category name, labels, or MVTec layout is required.
from tensor_md import (
LocationAwareTensorMahalanobisDetector,
PatchExtractionConfig,
load_image_patches,
load_normal_patches,
)
config = PatchExtractionConfig(
train_image_dir="data/normal",
image_size=(256, 256),
patch_size=(16, 16),
stride=16,
)
normal = load_normal_patches(config)
images = load_image_patches("data/to_check", config)
detector = LocationAwareTensorMahalanobisDetector(
patches_per_image=normal.patches_per_image,
)
detector.fit(normal)
scores = detector.score(images)
# Rows are images and columns are spatial patch locations.
score_maps = scores.reshape(len(images.image_paths), images.patches_per_image)
After fitting, the detector retains the inverse-square-root factors required
for scoring and releases the now-redundant covariance factors by default. Set
retain_covariances=True only when the original fitted covariance matrices are
needed for inspection or diagnostics.
Training and scoring images must use the same configuration. Location-aware modelling is most useful when images are approximately aligned, so the same grid location usually represents the same object part or texture region.
Optional score calibration
Raw tensor Mahalanobis scores already use the fitted mean and covariance. Optional z-score calibration runs the fitted detector on the normal training set and stores statistics of those scalar training scores. Its spatial scope is explicit and independent of mean or covariance shrinkage:
# One score mean and standard deviation per location.
detector = LocationAwareTensorMahalanobisDetector(
patches_per_image=normal.patches_per_image,
score_normalization="location_zscore",
)
# One score mean and standard deviation shared by all locations.
detector = LocationAwareTensorMahalanobisDetector(
patches_per_image=normal.patches_per_image,
score_normalization="global_zscore",
)
The available settings are:
"none": retain raw Mahalanobis scores; this is the default."location_zscore": calibrate each location separately. This intentionally adds location-specific calibration even when the fitted mean and covariance are fully shared."global_zscore": apply one calibration to every location. Because this is the same positive affine transformation for every score, it preserves score ordering and is mainly useful for a common threshold scale."zscore": backward-compatible alias for"location_zscore".
Calibration statistics are learned only from normal training scores. Test-set statistics are never used.
Using CNN features
Set input_representation="cnn_features" to model intermediate CNN features
instead of RGB patches. This example uses two built-in ResNet50 layers and
reduces both to 128 channels before stacking them as a tensor mode:
config = PatchExtractionConfig(
train_image_dir="data/normal",
input_representation="cnn_features",
cnn_backbone="ResNet50",
cnn_layer_names=("conv3_block4_out", "conv4_block6_out"),
cnn_dimensionality_reduction="pca",
cnn_reduction_dimensions=128,
cnn_feature_patch_size=(1, 1),
)
Available dimensionality-reduction settings are:
"pca": fit one channel PCA per selected layer using normal training descriptors."random": retain a reproducible random subset of channels from each layer."none"orNone: keep all channels.
Multiple feature maps must have the same retained channel count so they can be
stacked. Without reduction, their original channel counts must already match.
The loader raises a clear error if cnn_reduction_dimensions is larger than
the channel count of any selected layer.
Supplying any CNN
The tensor detector is not tied to ResNet. Pass a callable through
cnn_feature_extractor. It receives an NHWC float32 batch with values in
[0, 1] and returns either one NHWC feature-map batch or a list of them:
def extract_features(batch):
# Return shape: (N, H, W, C), or a list of such arrays.
return my_model(batch)
config = PatchExtractionConfig(
train_image_dir="data/normal",
input_representation="cnn_features",
cnn_feature_extractor=extract_features,
)
Keras and PyTorch models can also be wrapped with the convenience adapter:
from tensor_md import make_cnn_feature_extractor
extractor = make_cnn_feature_extractor(model, framework="pytorch")
config = PatchExtractionConfig(
train_image_dir="data/normal",
input_representation="cnn_features",
cnn_feature_extractor=extractor,
)
The adapter handles the common NCHW/NHWC layout conversion. A custom extractor remains appropriate for models with unusual inputs or outputs.
Neighbourhood scoring
The neighbourhood detector pools completed Mahalanobis scores across nearby grid locations. This can make localization less sensitive to small spatial movements:
from tensor_md import NeighborhoodScoreLocationAwareTensorMahalanobisDetector
detector = NeighborhoodScoreLocationAwareTensorMahalanobisDetector(
patches_per_image=normal.patches_per_image,
grid_shape=(16, 16),
score_neighbor_radius=1,
score_neighbor_pooling="weighted_mean",
)
detector.fit(normal)
scores = detector.score(images)
The product of grid_shape must equal patches_per_image.
Set score_neighbor_pooling="median" to suppress isolated score spikes while
retaining responses supported by several nearby grid locations. For example,
radius one applies a 3 x 3 median window. The available pooling modes are
"mean", "max", "median", and "weighted_mean".
Optional orientation-conditioned mean
For an elongated object whose position is stable but whose orientation changes, the detector can model the expected feature mean as a smooth function of the object angle. The covariance model remains location-specific and is fitted to the residuals after subtracting that conditional mean.
config = PatchExtractionConfig(
train_image_dir="data/normal",
test_image_dir="data/to_check",
input_representation="cnn_features",
image_context_mode="light_background_orientation",
)
datasets = load_patch_datasets(config)
detector = LocationAwareTensorMahalanobisDetector(
patches_per_image=datasets.train.patches_per_image,
conditioning="fourier_mean",
conditioning_order=4,
conditioning_ridge=1e-3,
)
detector.fit_dataset(datasets.train)
scores = detector.score_dataset(datasets.test)
Use dark_foreground_orientation for a bright object on a dark background.
For other geometries, supply image_context_extractor=callable; it receives an
image path and must return one finite scalar in radians. This option does not
rotate or crop images. It is intended only when the extracted angle has a clear,
consistent meaning for every image.
Score diagnostics
Diagnostics are optional and do not require anomaly labels. They compare the scores of the normal training images with the images being checked and can save arrays, distributions, heatmaps, or floating-point TIFF score maps:
artifacts = detector.fit_and_save_diagnostics(
normal,
images,
"outputs/diagnostics",
grid_shape=(16, 16),
formats=("npy", "json", "distribution", "heatmaps", "tiff"),
)
Available formats are npy, csv, json, tiff, distribution, and
heatmaps. Diagnostics are for inspection; benchmark metrics still require
the benchmark's official ground-truth masks and evaluator.
Saving a fitted detector
detector.save("models/detector.pkl")
restored = LocationAwareTensorMahalanobisDetector.load(
"models/detector.pkl"
)
scores = restored.score(images)
Model files use Python pickle and must only be loaded from a trusted source. Reuse the same image preprocessing and feature extractor when creating data for the restored detector.
MVTec AD layout
MVTec AD is optional. If the dataset is arranged as
<root>/<category>/train/good and <root>/<category>/test/..., both splits can
be loaded together:
from tensor_md import load_patch_datasets
config = PatchExtractionConfig(
category="bottle",
data_root="/path/to/mvtec",
input_representation="cnn_features",
cnn_backbone="ResNet50",
cnn_layer_names=("conv3_block4_out", "conv4_block6_out"),
cnn_dimensionality_reduction="pca",
cnn_reduction_dimensions=128,
)
datasets = load_patch_datasets(config)
The package does not download or bundle MVTec AD.
Release Notices
git pull --ff-only python scripts/release.py
License
MIT; see LICENSE.
Project details
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Filter files by name, interpreter, ABI, and platform.
If you're not sure about the file name format, learn more about wheel file names.
Copy a direct link to the current filters
File details
Details for the file tensor_md-0.1.8.tar.gz.
File metadata
- Download URL: tensor_md-0.1.8.tar.gz
- Upload date:
- Size: 54.4 kB
- Tags: Source
- Uploaded using Trusted Publishing? Yes
- Uploaded via: twine/6.1.0 CPython/3.13.14
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
a8ffed50da39bcc513699401ba273e536795aa1886a4649795f9bc1899776d11
|
|
| MD5 |
d315f7deafe39ec736f720b5babb7b43
|
|
| BLAKE2b-256 |
1795e46e4eef43afbd96c0d0ed7a3dfcd4cae61cd89f6dfae8202d1a3adc7c94
|
Provenance
The following attestation bundles were made for tensor_md-0.1.8.tar.gz:
Publisher:
publish.yml on greinald/tensor-md
-
Statement:
-
Statement type:
https://in-toto.io/Statement/v1 -
Predicate type:
https://docs.pypi.org/attestations/publish/v1 -
Subject name:
tensor_md-0.1.8.tar.gz -
Subject digest:
a8ffed50da39bcc513699401ba273e536795aa1886a4649795f9bc1899776d11 - Sigstore transparency entry: 2218883324
- Sigstore integration time:
-
Permalink:
greinald/tensor-md@1038c9b5ac67db300282f0ab18ba905adf37a299 -
Branch / Tag:
refs/tags/v0.1.8 - Owner: https://github.com/greinald
-
Access:
public
-
Token Issuer:
https://token.actions.githubusercontent.com -
Runner Environment:
github-hosted -
Publication workflow:
publish.yml@1038c9b5ac67db300282f0ab18ba905adf37a299 -
Trigger Event:
push
-
Statement type:
File details
Details for the file tensor_md-0.1.8-py3-none-any.whl.
File metadata
- Download URL: tensor_md-0.1.8-py3-none-any.whl
- Upload date:
- Size: 50.4 kB
- Tags: Python 3
- Uploaded using Trusted Publishing? Yes
- Uploaded via: twine/6.1.0 CPython/3.13.14
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
1c5746eca4638002f844c2ecfdc212a52cfb5f45d56b5293cc836da17772f8ca
|
|
| MD5 |
839a3d58cd2ff565ca482bf132350897
|
|
| BLAKE2b-256 |
01ce0d333754991b915b29316a6157c459b8317dd330aa7c919613e99be4ce84
|
Provenance
The following attestation bundles were made for tensor_md-0.1.8-py3-none-any.whl:
Publisher:
publish.yml on greinald/tensor-md
-
Statement:
-
Statement type:
https://in-toto.io/Statement/v1 -
Predicate type:
https://docs.pypi.org/attestations/publish/v1 -
Subject name:
tensor_md-0.1.8-py3-none-any.whl -
Subject digest:
1c5746eca4638002f844c2ecfdc212a52cfb5f45d56b5293cc836da17772f8ca - Sigstore transparency entry: 2218883394
- Sigstore integration time:
-
Permalink:
greinald/tensor-md@1038c9b5ac67db300282f0ab18ba905adf37a299 -
Branch / Tag:
refs/tags/v0.1.8 - Owner: https://github.com/greinald
-
Access:
public
-
Token Issuer:
https://token.actions.githubusercontent.com -
Runner Environment:
github-hosted -
Publication workflow:
publish.yml@1038c9b5ac67db300282f0ab18ba905adf37a299 -
Trigger Event:
push
-
Statement type: