Presidio image redactor package

These details have not been verified by PyPI

Project links

Homepage

License
- OSI Approved :: MIT License
Natural Language
- English
Operating System
- OS Independent
Programming Language

Project description

Presidio Image Redactor

Please notice, this package is still in alpha and not production ready.

Description

The Presidio Image Redactor is a Python based module for detecting and redacting PII text entities in images.

Deploy Presidio image redactor to Azure

Use the following button to deploy presidio image redactor to your Azure subscription.

Image Redactor Design

Installation

Pre-requisites:

Install Tesseract OCR by following the instructions on how to install it for your operating system.

For now, image redactor only supports version 4.0.0

As package:

To get started with Presidio-image-redactor, run the following:

pip install presidio-image-redactor

Once Installed, run the following command to download the default spacy model needed for Presidio Analyzer:

python -m spacy download en_core_web_lg

Getting started

The engine will receive 2 parameters:

Image to redact.
Color fill to redact with, by default color fill will be black. Can either be an int or tuple (0,0,0)

from PIL import Image
from presidio_image_redactor import ImageRedactorEngine

# Get the image to redact using PIL lib (pillow)
image = Image.open("ocr_text.png")

# Initialize the engine
engine = ImageRedactorEngine()

# Redact the image with pink color
redacted_image = engine.redact(image, (255, 192, 203))

# save the redacted image 
redacted_image.save("new_image.png")
# open the image for viewing
redacted_image.show()

As docker service:

In folder presidio/presidio-image-redactor run:

docker-compose up -d

HTTP API

redact

Receives an image and color fill (optional, default is black). Redact the image PII text and returns a new redacted image.

POST /redact

Payload:

Sent as multipart-form. Contains image file and data of the required color fill.

{
  "data": "{'color_fill':'0,0,0'}"
}

Result:

200 OK

curl example:

# use ocr_test.png as the image to redact, and 255 as the color fill. 
# out.png is the new redacted image received from the server.
curl -XPOST "http://localhost:3000/redact" -H "content-type: multipart/form-data" -F "image=@ocr_test.png" -F "data=\"{'color_fill':'255'}\"" > out.png

Python script example can be found under: /presidio/e2e-tests/tests/test_image_redactor.py

Project details

These details have not been verified by PyPI

Project links

Homepage

License
- OSI Approved :: MIT License
Natural Language
- English
Operating System
- OS Independent
Programming Language

Release history Release notifications | RSS feed

0.0.53

Jul 22, 2024

0.0.52

Mar 29, 2024

0.0.51

Feb 12, 2024

0.0.50

Jan 22, 2024

0.0.49

Oct 30, 2023

0.0.48

Jun 21, 2023

0.0.47

Jun 4, 2023

0.0.46

Jan 31, 2023

0.0.45

Dec 14, 2022

0.0.44

May 8, 2022

0.0.43

Jan 26, 2022

0.0.42

Oct 3, 2021

0.0.41

Jun 14, 2021

0.0.4

May 10, 2021

This version

0.0.3

Apr 12, 2021

0.0.2

Apr 12, 2021

0.0.1

Mar 1, 2021

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distributions

No source distribution files available for this release.See tutorial on generating distribution archives.

Built Distribution

presidio_image_redactor-0.0.3-py3-none-any.whl (8.8 kB view details)

Uploaded Apr 12, 2021 Python 3

File details

Details for the file presidio_image_redactor-0.0.3-py3-none-any.whl.

File metadata

Download URL: presidio_image_redactor-0.0.3-py3-none-any.whl
Upload date: Apr 12, 2021
Size: 8.8 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/3.4.1 importlib_metadata/3.10.0 pkginfo/1.7.0 requests/2.25.1 requests-toolbelt/0.9.1 tqdm/4.60.0 CPython/3.8.8

File hashes

Hashes for presidio_image_redactor-0.0.3-py3-none-any.whl
Algorithm	Hash digest
SHA256	`74264909de7705a5a97b5e608f23df96a83cb4e84c647ea3816d3b318873b67d`
MD5	`59eb76d9f9e87ce50acd288e1f658bd0`
BLAKE2b-256	`f3e090e806451b9f3bfd4f1d4d7748a3b7d680da8a1f1960d601eeb67ea2d9d1`