A vision library for performing sliced inference on large images/small objects

These details have not been verified by PyPI

Project links

Homepage

Project description

SAHI: Slicing Aided Hyper Inference

A vision library for performing sliced inference on large images/small objects.

Overview

Object detection and instance segmentation are by far the most important fields of applications in Computer Vision. However, detection of small objects and inference on large images are still major issues in practical usage. Here comes the SAHI to help developers overcome these real-world problems.

Getting Started

Blogpost

Check the official SAHI blog post.

Installation

Install sahi using pip:

pip install sahi

On Windows, Shapely needs to be installed via Conda:

conda install -c conda-forge shapely

Install your desired version of pytorch and torchvision:

pip install torch torchvision

Install your desired detection framework (such as mmdet or yolov5):

pip install mmdet mmcv

pip install yolov5

Usage

From Python:

Sliced inference:

result = get_sliced_prediction(
    image,
    detection_model,
    slice_height = 256,
    slice_width = 256,
    overlap_height_ratio = 0.2,
    overlap_width_ratio = 0.2
)

Check YOLOv5 + SAHI demo:

Check MMDetection + SAHI demo:

Slice an image:

from sahi.slicing import slice_image

slice_image_result, num_total_invalid_segmentation = slice_image(
    image=image_path,
    output_file_name=output_file_name,
    output_dir=output_dir,
    slice_height=256,
    slice_width=256,
    overlap_height_ratio=0.2,
    overlap_width_ratio=0.2,
)

Slice a coco formatted dataset:

from sahi.slicing import slice_coco

coco_dict, coco_path = slice_coco(
    coco_annotation_file_path=coco_annotation_file_path,
    image_dir=image_dir,
    slice_height=256,
    slice_width=256,
    overlap_height_ratio=0.2,
    overlap_width_ratio=0.2,
)

Refer to slicing notebook for detailed usage.

From CLI:

python scripts/predict.py --source image/file/or/folder --model_path path/to/model --config_path path/to/config

will perform sliced inference on default parameters and export the prediction visuals to runs/predict/exp folder.

You can specify sliced inference parameters as:

python scripts/predict.py --slice_width 256 --slice_height 256 --overlap_height_ratio 0.1 --overlap_width_ratio 0.1 --iou_thresh 0.25 --source image/file/or/folder --model_path path/to/model --config_path path/to/config

Specify postprocess type as --postprocess_type UNIONMERGE or --postprocess_type NMS to be applied over sliced predictions
Specify postprocess match metric as --match_metric IOS for intersection over smaller area or --match_metric IOU for intersection over union
Specify postprocess match threshold as --match_thresh 0.5
Add --class_agnostic argument to ignore category ids of the predictions during postprocess (merging/nms)
If you want to export prediction pickles and cropped predictions add --pickle and --crop arguments. If you want to change crop extension type, set it as --visual_export_format JPG.
If you don't want to export prediction visuals, add --novisual argument.
If you want to perform standard prediction instead of sliced prediction, add --standard_pred argument.

python scripts/predict.py --coco_file path/to/coco/file --source coco/images/directory --model_path path/to/model --config_path path/to/config

will perform inference using provided coco file, then export results as a coco json file to runs/predict/exp/results.json

Find detailed info on script usage (predict, coco2yolov5, coco_error_analysis) at SCRIPTS.md.

COCO Utilities

COCO dataset creation:

import required classes:

from sahi.utils.coco import Coco, CocoCategory, CocoImage, CocoAnnotation

init Coco object:

coco = Coco()

add categories starting from id 0:

coco.add_category(CocoCategory(id=0, name='human'))
coco.add_category(CocoCategory(id=1, name='vehicle'))

create a coco image:

coco_image = CocoImage(file_name="image1.jpg", height=1080, width=1920)

add annotations to coco image:

coco_image.add_annotation(
  CocoAnnotation(
    bbox=[x_min, y_min, width, height],
    category_id=0,
    category_name='human'
  )
)
coco_image.add_annotation(
  CocoAnnotation(
    bbox=[x_min, y_min, width, height],
    category_id=1,
    category_name='vehicle'
  )
)

add coco image to Coco object:

coco.add_image(coco_image)

after adding all images, convert coco object to coco json:

coco_json = coco.json

you can export it as json file:

from sahi.utils.file import save_json

save_json(coco_json, "coco_dataset.json")

Convert COCO dataset to ultralytics/yolov5 format:

from sahi.utils.coco import Coco

# init Coco object
coco = Coco.from_coco_dict_or_path("coco.json", image_dir="coco_images/")

# export converted YoloV5 formatted dataset into given output_dir with a 85% train/15% val split
coco.export_as_yolov5(
  output_dir="output/folder/dir",
  train_split_rate=0.85
)

Get dataset stats:

from sahi.utils.coco import Coco

# init Coco object
coco = Coco.from_coco_dict_or_path("coco.json")

# get dataset stats
coco.stats
{
  'num_images': 6471,
  'num_annotations': 343204,
  'num_categories': 2,
  'num_negative_images': 0,
  'num_images_per_category': {'human': 5684, 'vehicle': 6323},
  'num_annotations_per_category': {'human': 106396, 'vehicle': 236808},
  'min_num_annotations_in_image': 1,
  'max_num_annotations_in_image': 902,
  'avg_num_annotations_in_image': 53.037243084530985,
  'min_annotation_area': 3,
  'max_annotation_area': 328640,
  'avg_annotation_area': 2448.405738278109,
  'min_annotation_area_per_category': {'human': 3, 'vehicle': 3},
  'max_annotation_area_per_category': {'human': 72670, 'vehicle': 328640},
}

Find detailed info on COCO utilities (yolov5 conversion, slicing, subsampling, filtering, merging, splitting) at COCO.md.

MOT Challenge Utilities

MOT Challenge formatted ground truth dataset creation:

import required classes:

from sahi.utils.mot import MotAnnotation, MotFrame, MotVideo

init video:

mot_video = MotVideo(name="sequence_name")

init first frame:

mot_frame = MotFrame()

add annotations to frame:

mot_frame.add_annotation(
  MotAnnotation(bbox=[x_min, y_min, width, height])
)

mot_frame.add_annotation(
  MotAnnotation(bbox=[x_min, y_min, width, height])
)

add frame to video:

mot_video.add_frame(mot_frame)

export in MOT challenge format:

mot_video.export(export_dir="mot_gt", type="gt")

your MOT challenge formatted ground truth files are ready under mot_gt/sequence_name/ folder.

Find detailed info on MOT utilities (ground truth dataset creation, exporting tracker metrics in mot challenge format) at MOT.md.

Contributing

sahi library currently supports all YOLOv5 models and MMDetection models. Moreover, it is easy to add new frameworks.

All you need to do is, creating a new class in model.py that implements DetectionModel class. You can take the MMDetection wrapper or YOLOv5 wrapper as a reference.

Contributers

Fatih Cagatay Akyon

Cemil Cengiz

Sinan Onur Altinuc

Project details

These details have not been verified by PyPI

Project links

Homepage

Release history Release notifications | RSS feed

0.11.18

Jul 10, 2024

0.11.16

May 20, 2024

0.11.15

Nov 5, 2023

0.11.14

May 15, 2023

0.11.13

Mar 24, 2023

0.11.12

Mar 7, 2023

0.11.11

Jan 15, 2023

0.11.10

Jan 5, 2023

0.11.9

Dec 24, 2022

0.11.8

Dec 23, 2022

0.11.7

Dec 20, 2022

0.11.6

Dec 1, 2022

0.11.5

Nov 27, 2022

0.11.4

Nov 15, 2022

0.11.3

Nov 14, 2022

0.11.2

Nov 14, 2022

0.11.1

Nov 1, 2022

0.11.0

Oct 29, 2022

0.10.8

Oct 26, 2022

0.10.7

Sep 27, 2022

0.10.6

Sep 25, 2022

0.10.5

Sep 4, 2022

0.10.4

Aug 12, 2022

0.10.3

Aug 2, 2022

0.10.2

Jul 28, 2022

0.10.1

Jun 25, 2022

0.10.0

Jun 21, 2022

0.9.4

May 28, 2022

0.9.3

May 8, 2022

0.9.2

Apr 9, 2022

0.9.1

Mar 17, 2022

0.9.0

Feb 12, 2022

0.8.22

Jan 13, 2022

0.8.21

Jan 12, 2022

0.8.20

Jan 11, 2022

0.8.19

Jan 6, 2022

0.8.18

Jan 2, 2022

0.8.16

Dec 26, 2021

0.8.15

Dec 15, 2021

0.8.14

Dec 11, 2021

0.8.13

Dec 6, 2021

0.8.12

Nov 30, 2021

0.8.11

Nov 24, 2021

0.8.10

Nov 23, 2021

0.8.9

Nov 15, 2021

0.8.8

Nov 2, 2021

0.8.6

Oct 19, 2021

0.8.5

Oct 19, 2021

0.8.4

Sep 13, 2021

0.8.3

Sep 10, 2021

0.8.2

Sep 9, 2021

0.8.1

Sep 8, 2021

0.8.0

Sep 2, 2021

0.7.4

Aug 24, 2021

0.7.3

Aug 8, 2021

0.7.2

Aug 3, 2021

0.7.1

Jul 28, 2021

0.7.0

Jul 13, 2021

0.6.2

Jul 13, 2021

0.6.1

Jul 9, 2021

0.6.0

Jul 6, 2021

0.5.2

Jul 5, 2021

0.5.1

Jul 3, 2021

0.5.0

Jun 28, 2021

0.4.8

Jun 23, 2021

0.4.7

Jun 23, 2021

This version

0.4.6

Jun 14, 2021

0.4.5

Jun 12, 2021

0.4.4

Jun 10, 2021

0.4.3

May 31, 2021

0.4.2

May 22, 2021

0.4.1

May 22, 2021

0.4.0

May 22, 2021

0.3.19

May 15, 2021

0.3.18

May 10, 2021

0.3.17

May 8, 2021

0.3.16

May 8, 2021

0.3.15

May 7, 2021

0.3.14

May 7, 2021

0.3.13

May 6, 2021

0.3.12

May 4, 2021

0.3.11

May 1, 2021

0.3.10

May 1, 2021

0.3.9

May 1, 2021

0.3.8

Apr 28, 2021

0.3.7

Apr 28, 2021

0.3.6

Apr 27, 2021

0.3.5

Apr 27, 2021

0.3.4

Apr 24, 2021

0.3.3

Apr 11, 2021

0.3.2

Apr 11, 2021

0.3.1

Feb 26, 2021

0.3.0

Feb 23, 2021

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

sahi-0.4.6.tar.gz (51.4 kB view details)

Uploaded Jun 14, 2021 Source

Built Distribution

sahi-0.4.6-py3-none-any.whl (54.6 kB view details)

Uploaded Jun 14, 2021 Python 3

File details

Details for the file sahi-0.4.6.tar.gz.

File metadata

Download URL: sahi-0.4.6.tar.gz
Upload date: Jun 14, 2021
Size: 51.4 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/3.4.1 importlib_metadata/4.5.0 pkginfo/1.7.0 requests/2.25.1 requests-toolbelt/0.9.1 tqdm/4.61.1 CPython/3.9.5

File hashes

Hashes for sahi-0.4.6.tar.gz
Algorithm	Hash digest
SHA256	`c03083ad3cdab325e09734f3b89363a380b3fe75b6d98ad27d0477e9fcf2b934`
MD5	`efd37dc3333866f479b1e9e5bae5c56b`
BLAKE2b-256	`bdbd93a179f825c0732641a8ef4ed3416481b79180985bfee91e90311037fb15`

See more details on using hashes here.

File details

Details for the file sahi-0.4.6-py3-none-any.whl.

File metadata

Download URL: sahi-0.4.6-py3-none-any.whl
Upload date: Jun 14, 2021
Size: 54.6 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/3.4.1 importlib_metadata/4.5.0 pkginfo/1.7.0 requests/2.25.1 requests-toolbelt/0.9.1 tqdm/4.61.1 CPython/3.9.5

File hashes

Hashes for sahi-0.4.6-py3-none-any.whl
Algorithm	Hash digest
SHA256	`e62e83f58d75ae004044a5aeced124825d3b9dac307849af71e90e3d09945d15`
MD5	`81f6ae1adb433437a20ac1c84d958e75`
BLAKE2b-256	`e42fe56fcc9baaa3531387005cfe984a7ceb44c095bab90c5418de9702098124`